Proveedor · 2026-09-28
fireworks-ai
| Modelo | Entrada / 1M | Salida / 1M | Contexto | Proveedores | Etiquetas |
|---|---|---|---|---|---|
| GPT OSS 120B | $0.150 | $0.600 | 131K | 1 | tools · reasoning · open-weights |
| Inkling | $1.00 | $4.05 | 1.05M | 1 | tools · reasoning · vision · open-weights |
| Ember-1 | $3.00 | $15.00 | 1.05M | 1 | tools · reasoning · vision |
| Nemotron 3.5 Lightning 30B A3Bderivado | $0.050 | $0.200 | 262K | 1 | tools · json · reasoning · open-weights |
| GLM 5.3 Flashderivado | $0.150 | $0.500 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
| DeepSeek V4.1 Flashderivado | $0.220 | $0.660 | 1M | 1 | tools · json · reasoning · vision · open-weights |
| MiniMax Latestderivado | $0.300 | $1.20 | 512K | 1 | tools · reasoning · open-weights |
| Nemotron 3 Ultra 550B A55Bderivado | $0.600 | $2.40 | 262K | 1 | tools · reasoning · open-weights |
| GLM 5.3derivado | $1.40 | $4.40 | 1.05M | 1 | tools · json · reasoning · open-weights |
| Qwen3.8 2.4T A95Bderivado | $2.00 | $6.00 | 262K | 1 | tools · json · reasoning · open-weights |
| Qwen3.8 Maxderivado | $2.00 | $6.00 | 262K | 1 | tools · reasoning · vision |
| Qwen Max Latest (Qwen3.8 Max)derivado | $2.00 | $6.00 | 262K | 1 | tools · reasoning · vision |
| GLM 5.3 Fast (Latest)derivado | $2.10 | $6.60 | 1.05M | 1 | tools · json · reasoning · open-weights |
| GLM 5.3 Fastderivado | $2.10 | $6.60 | 1.05M | 1 | tools · json · reasoning · open-weights |
| Kimi Fast Latestderivado | $4.50 | $22.50 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
Frequently asked questions
How many AI models does fireworks-ai offer?
We track 3 canonical fireworks-ai models plus 12 community fine-tunes / derivatives (excluded from the main table). The list is recomputed daily.
Which fireworks-ai model is the cheapest?
GPT OSS 120B is currently the lowest-priced fireworks-ai model, at $0.150 per 1M input tokens and $0.600 per 1M output tokens. For the full apples-to-apples list, see /pricing/cheapest-llm-api.
Which fireworks-ai model has the largest context window?
Inkling leads at 1.05M tokens. This is the total of prompt + completion.
Which fireworks-ai models support tool calling?
Multiple fireworks-ai models support tool calling, with GPT OSS 120B being a popular pick. The capability column in the table above marks every model with fireworks-ai tool-calling support.
Which fireworks-ai models accept image input?
Inkling accepts image input. Other vision-capable fireworks-ai models are tagged 'vision' in the table above. See /capabilities/vision for a cross-vendor comparison.
What are the best alternatives to fireworks-ai?
Depends on the use case. For raw cost savings, look at /pricing/cheapest-llm-api. For agent-oriented workloads, /best/best-ai-model-for-agents. For long-document workflows, /best/best-long-context-llm.
How fresh is this fireworks-ai pricing data?
Daily. Our pipeline syncs every morning and rebuilds these pages on data change, so list-price moves and new model releases land within roughly 24 hours.
Explore more
Top fireworks-ai models
- GPT OSS 120B$0.15 in / $0.60 out
- Inkling$1.00 in / $4.05 out
- Ember-1$3.00 in / $15.00 out
Browse by use case
Browse by capability
Última actualización:
Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.
Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.