OpenAI compatible API · Attested · Public status

Fireworks AI

Explore Fireworks AI models on TrustedRouter with current routes, token pricing, policy sources, privacy posture, regional availability, and API support.

Verify gateway
Onebase URL to migrate
100sof models and routes
0prompt or output logs. Always.

Fireworks AIfireworks

No provider claim

All providers

ProviderFireworks AI
Provider websitehttps://fireworks.ai/
Models11 public models
Prepaid routes11
BYOK routes11
Zero data retentionnot claimed
Confidential computenot claimed
Provider E2EEnot claimed
Policy noteNo provider-ZDR claim is tracked here. Fireworks publishes security, privacy, and zero-retention documentation; enable a contracted ZDR posture before marking this provider as ZDR.
Policy source

Measured performance

71 samples

Continuously sampled across Fireworks AI's routed models: p50 TTFT, effective throughput, and success rate. Effective throughput uses provider-reported output tokens over complete request time. Unsupported route and probe-configuration rows are separated from provider downtime. No prompt or output content stored.

p50 TTFT1867 ms
Effective throughput49 tok/s n=14
Uptime98.59%
Modelp50 TTFTp50 TTFBEffective throughputUptimeConfig excludedAvailability samples
moonshotai/kimi-k2.6 881 ms 881 ms 46 tok/s n=2 100.00% 8
openai/gpt-oss-20b 1081 ms 1081 ms 100.00% 5
z-ai/glm-5.2-fast 1258 ms 1258 ms 59 tok/s n=2 100.00% 6
openai/gpt-oss-120b 1509 ms 1509 ms 98 tok/s n=2 100.00% 5
deepseek/deepseek-v4-pro 1867 ms 1867 ms 36 tok/s n=1 100.00% 12
z-ai/glm-5.2 2259 ms 2259 ms 36 tok/s n=1 100.00% 6
moonshotai/kimi-k3 2571 ms 2571 ms 44 tok/s n=1 100.00% 7
deepseek/deepseek-v4-flash-0731 2891 ms 2891 ms 44 tok/s n=2 100.00% 9
moonshotai/kimi-k2.7-code 3749 ms 3749 ms 51 tok/s n=1 100.00% 8
deepseek/deepseek-v4-flash 4038 ms 4037 ms 63 tok/s n=1 100.00% 1
minimax/minimax-m3 3921 ms 3920 ms 67 tok/s n=1 75.00% 4

Fireworks AI performance history · Full provider & model leaderboard.

Provider models

Models served by Fireworks AI.

Each row links to pricing, provider, benchmark, and API pages for the model.

Model AI IQ Context Endpoints Prompt Completion Routes
deepseek/deepseek-v4-flash
DeepSeek: DeepSeek V4 Flash 0423
IQ 108#46 1,048,576 2 $0.147/1M $0.294/1M prepaid BYOK
deepseek/deepseek-v4-flash-0731
DeepSeek: DeepSeek V4 Flash 0731
1,048,576 2 $0.147/1M $0.294/1M prepaid BYOK
deepseek/deepseek-v4-pro
DeepSeek: DeepSeek V4 Pro
IQ 115#30 1,048,576 2 $1.827/1M $3.654/1M prepaid BYOK
minimax/minimax-m3
MiniMax: MiniMax M3
IQ 114#33 1,048,576 2 $0.315/1M $1.26/1M prepaid BYOK
moonshotai/kimi-k2.6
MoonshotAI: Kimi K2.6
IQ 119#19 262,144 2 $0.9975/1M $4.2/1M prepaid BYOK
moonshotai/kimi-k2.7-code
MoonshotAI: Kimi K2.7 Code
IQ 118#22 262,144 2 $0.9975/1M $4.2/1M prepaid BYOK
moonshotai/kimi-k3
MoonshotAI: Kimi K3
IQ 123#14 1,048,576 2 $3.15/1M $15.75/1M prepaid BYOK
openai/gpt-oss-120b
OpenAI: gpt-oss-120b
IQ 105#53 131,072 2 $0.1575/1M $0.63/1M prepaid BYOK
openai/gpt-oss-20b
OpenAI: gpt-oss-20b
IQ 100#70 131,072 2 $0.0735/1M $0.315/1M prepaid BYOK
z-ai/glm-5.2
Z.ai: GLM 5.2
IQ 120#16 1,048,576 2 $1.47/1M $4.62/1M prepaid BYOK
z-ai/glm-5.2-fast
GLM 5.2 Fast on Fireworks
1,048,576 2 $2.94/1M $9.24/1M prepaid BYOK
Workspace access

Sign in

Choose a sign in method to access your TrustedRouter workspace.

By signing in you agree to the terms of service and privacy policy.