KI‑Modell‑Intelligenz

Anbieter · 2026-09-28

fireworks-ai

3 kanonische Modelle15 Einträge insgesamt (inkl. Derivate)
ModellEingabe / 1MAusgabe / 1MKontextHosterTags
GPT OSS 120B$0.150$0.600131K1tools · reasoning · open-weights
Inkling$1.00$4.051.05M1tools · reasoning · vision · open-weights
Ember-1$3.00$15.001.05M1tools · reasoning · vision
Nemotron 3.5 Lightning 30B A3BDerivat$0.050$0.200262K1tools · json · reasoning · open-weights
GLM 5.3 FlashDerivat$0.150$0.5001.05M1tools · json · reasoning · vision · open-weights
DeepSeek V4.1 FlashDerivat$0.220$0.6601M1tools · json · reasoning · vision · open-weights
MiniMax LatestDerivat$0.300$1.20512K1tools · reasoning · open-weights
Nemotron 3 Ultra 550B A55BDerivat$0.600$2.40262K1tools · reasoning · open-weights
GLM 5.3Derivat$1.40$4.401.05M1tools · json · reasoning · open-weights
Qwen3.8 2.4T A95BDerivat$2.00$6.00262K1tools · json · reasoning · open-weights
Qwen3.8 MaxDerivat$2.00$6.00262K1tools · reasoning · vision
Qwen Max Latest (Qwen3.8 Max)Derivat$2.00$6.00262K1tools · reasoning · vision
GLM 5.3 Fast (Latest)Derivat$2.10$6.601.05M1tools · json · reasoning · open-weights
GLM 5.3 FastDerivat$2.10$6.601.05M1tools · json · reasoning · open-weights
Kimi Fast LatestDerivat$4.50$22.501.05M1tools · json · reasoning · vision · open-weights

Frequently asked questions

How many AI models does fireworks-ai offer?

We track 3 canonical fireworks-ai models plus 12 community fine-tunes / derivatives (excluded from the main table). The list is recomputed daily.

Which fireworks-ai model is the cheapest?

GPT OSS 120B is currently the lowest-priced fireworks-ai model, at $0.150 per 1M input tokens and $0.600 per 1M output tokens. For the full apples-to-apples list, see /pricing/cheapest-llm-api.

Which fireworks-ai model has the largest context window?

Inkling leads at 1.05M tokens. This is the total of prompt + completion.

Which fireworks-ai models support tool calling?

Multiple fireworks-ai models support tool calling, with GPT OSS 120B being a popular pick. The capability column in the table above marks every model with fireworks-ai tool-calling support.

Which fireworks-ai models accept image input?

Inkling accepts image input. Other vision-capable fireworks-ai models are tagged 'vision' in the table above. See /capabilities/vision for a cross-vendor comparison.

What are the best alternatives to fireworks-ai?

Depends on the use case. For raw cost savings, look at /pricing/cheapest-llm-api. For agent-oriented workloads, /best/best-ai-model-for-agents. For long-document workflows, /best/best-long-context-llm.

How fresh is this fireworks-ai pricing data?

Daily. Our pipeline syncs every morning and rebuilds these pages on data change, so list-price moves and new model releases land within roughly 24 hours.

Zuletzt aktualisiert:

Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.

Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.