Provider · 2026-10-04
llmgateway-providers
| Model | Input / 1M | Output / 1M | Context | Providers | Tags |
|---|---|---|---|---|---|
| GPT OSS 20B (Consensus Protocol) | $0.040 | $0.190 | 128K | 2 | tools · reasoning · open-weights |
| InclusionAI Ling 3.0 Flash (DeepInfra) | $0.060 | $0.180 | 262K | 1 | tools · reasoning |
| Ling 3.0 Flash VL (DeepInfra) | $0.060 | $0.180 | 131K | 1 | tools · json · reasoning · vision |
| Muse Spark 1.2 Contributor (Meta Contributor) | $0.100 | $0.200 | 1.05M | 1 | tools · json · reasoning · vision |
| Muse Spark 1.3 Contributor (Meta Contributor) | $0.100 | $0.200 | 1.05M | 1 | tools · json · reasoning · vision |
| Granite 4.2 8B (DeepInfra) | $0.060 | $0.250 | 131K | 1 | tools · reasoning |
| Hy-MT2 Plus (Tencent Cloud) | $0.074 | $0.295 | 8K | 1 | — |
| MiMo V2.6 Flash (DeepInfra) | $0.140 | $0.280 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
| Hy3 (DeepInfra) | $0.140 | $0.580 | 262K | 2 | tools · json · reasoning · open-weights |
| GPT OSS 120B (Groq) | $0.150 | $0.750 | 131K | 6 | tools · reasoning · open-weights |
| MiMo V2.6 Pro (DeepInfra) | $0.435 | $0.870 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
| MiMo V2.5 Pro (Tencent Cloud) | $0.435 | $0.870 | 1M | 1 | tools · reasoning · open-weights |
| Muse Glimmer 30B (DeepInfra) | $0.300 | $1.20 | 131K | 2 | tools · json · reasoning · vision · open-weights |
| Hy4 Preview (Tencent Cloud) | $0.834 | $2.50 | 1M | 1 | tools · json · reasoning · open-weights |
| Inkling (DeepInfra) | $0.950 | $4.05 | 524K | 1 | reasoning · vision · open-weights |
| Fugu Max (Sakana AI) | $2.00 | $6.00 | 1M | 1 | tools · json · reasoning · vision |
| Fugu Ultra v2.0 (Sakana AI) | $5.00 | $30.00 | 1M | 1 | tools · json · reasoning · vision |
| Fugu Ultra (Sakana AI) | $5.00 | $30.00 | 1M | 1 | tools · reasoning · vision |
| Atria Dawn Preview (Atria) | Unknown | Unknown | 262K | 1 | tools · reasoning |
| GLM-4 32B (0414-128k) (Z AI)derivative | $0.100 | $0.100 | 128K | 1 | tools |
| Seed 1.6 Flash (250715) (ByteDance)derivative | $0.070 | $0.300 | 256K | 1 | tools · reasoning · vision |
| GLM-4.6V FlashX (Z AI)derivative | $0.040 | $0.400 | 128K | 1 | tools · reasoning · vision |
| Qwen3 VL Flash (Alibaba Cloud)derivative | $0.050 | $0.400 | 262K | 1 | tools · vision |
| DeepSeek V4 Flash (Alibaba Cloud)derivative | $0.200 | $0.400 | 1M | 2 | tools · reasoning · open-weights |
| DeepSeek V3.2 (ByteDance)derivative | $0.280 | $0.420 | 131K | 1 | tools · reasoning · open-weights |
| MiniMax Text 01 (MiniMax)derivative | $0.200 | $1.10 | 1M | 1 | tools · reasoning |
| DeepSeek V4.1 Flash (Alibaba Cloud)derivative | $0.300 | $1.20 | 1M | 1 | tools · reasoning · vision · open-weights |
| Qwen Coder Plus (Alibaba Cloud)derivative | $0.502 | $1.00 | 131K | 1 | tools |
| Qwen Plus Latest (Alibaba Cloud)derivative | $0.400 | $1.20 | 1M | 1 | tools |
| Qwen3 235B A22B Instruct 2507 (Cerebras)derivative | $0.600 | $1.20 | 262K | 1 | tools · open-weights |
| Llama 3.3 70B Instruct (Cerebras)derivative | $0.850 | $1.20 | 128K | 1 | tools · open-weights |
| Seed 1.6 (250915) (ByteDance)derivative | $0.250 | $2.00 | 256K | 1 | tools · reasoning · vision |
| Seed 1.8 (251228) (ByteDance)derivative | $0.250 | $2.00 | 256K | 1 | tools · reasoning · vision |
| Seed 1.6 (250615) (ByteDance)derivative | $0.250 | $2.00 | 256K | 1 | tools · reasoning · vision |
| Gemma 4 31B IT (Cerebras)derivative | $0.990 | $1.49 | 131K | 1 | tools · reasoning · vision · open-weights |
| GLM-5 (Alibaba Cloud)derivative | $0.573 | $2.58 | 203K | 1 | tools · reasoning · open-weights |
| Seed 2.0 Code Preview (260328) (ByteDance)derivative | $0.500 | $3.00 | 262K | 1 | tools · reasoning · vision |
| Seed 2.0 Pro (260328) (ByteDance)derivative | $0.500 | $3.00 | 262K | 1 | tools · json · reasoning · vision |
| Kimi K2.5 (Alibaba Cloud)derivative | $0.574 | $3.01 | 262K | 1 | tools · reasoning · vision · open-weights |
| Qwen3.5 397B A17B (Alibaba Cloud)derivative | $0.600 | $3.60 | 262K | 1 | tools · reasoning · vision |
| GLM-4.7 (Cerebras)derivative | $2.25 | $2.75 | 200K | 2 | tools · reasoning · open-weights |
| GLM-4.5 AirX (Z AI)derivative | $1.10 | $4.50 | 128K | 1 | tools |
| GLM-5.2 (Alibaba Cloud)derivative | $1.40 | $4.40 | 1M | 2 | tools · reasoning · open-weights |
| GLM-5.3 (Alibaba Cloud)derivative | $1.40 | $4.40 | 1.05M | 2 | tools · reasoning · open-weights |
| DeepSeek V4 Pro (Alibaba Cloud)derivative | $2.40 | $4.80 | 1M | 2 | tools · reasoning · open-weights |
| Grok 4.20 Beta Non-Reasoning (0309) (xAI)derivative | $2.00 | $6.00 | 2M | 1 | tools · vision |
| GLM-4.5 X (Z AI)derivative | $2.20 | $8.90 | 128K | 1 | tools · reasoning |
| Kimi K3 (Alibaba Cloud)derivative | $3.00 | $15.00 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
Frequently asked questions
How many AI models does llmgateway-providers offer?
We track 19 canonical llmgateway-providers models plus 29 community fine-tunes / derivatives (excluded from the main table). The list is recomputed daily.
Which llmgateway-providers model is the cheapest?
GPT OSS 20B (Consensus Protocol) is currently the lowest-priced llmgateway-providers model, at $0.040 per 1M input tokens and $0.190 per 1M output tokens. For the full apples-to-apples list, see /pricing/cheapest-llm-api.
Which llmgateway-providers model has the largest context window?
Muse Spark 1.2 Contributor (Meta Contributor) leads at 1.05M tokens. This is the total of prompt + completion.
Which llmgateway-providers models support tool calling?
Multiple llmgateway-providers models support tool calling, with GPT OSS 20B (Consensus Protocol) being a popular pick. The capability column in the table above marks every model with llmgateway-providers tool-calling support.
Which llmgateway-providers models accept image input?
Ling 3.0 Flash VL (DeepInfra) accepts image input. Other vision-capable llmgateway-providers models are tagged 'vision' in the table above. See /capabilities/vision for a cross-vendor comparison.
What are the best alternatives to llmgateway-providers?
Depends on the use case. For raw cost savings, look at /pricing/cheapest-llm-api. For agent-oriented workloads, /best/best-ai-model-for-agents. For long-document workflows, /best/best-long-context-llm.
How fresh is this llmgateway-providers pricing data?
Daily. Our pipeline syncs every morning and rebuilds these pages on data change, so list-price moves and new model releases land within roughly 24 hours.
Explore more
Top llmgateway-providers models
- GPT OSS 120B (Groq)$0.15 in / $0.75 out
- GPT OSS 20B (Consensus Protocol)$0.04 in / $0.19 out
- Hy3 (DeepInfra)$0.14 in / $0.58 out
- Muse Glimmer 30B (DeepInfra)$0.30 in / $1.20 out
- InclusionAI Ling 3.0 Flash (DeepInfra)$0.06 in / $0.18 out
Browse by use case
Browse by capability
Last updated:
Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.
Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.