Provider · 2026-08-14
venice
| Model | Input / 1M | Output / 1M | Context | Providers | Tags |
|---|---|---|---|---|---|
| Venice Uncensored 1.2 | $0.200 | $0.900 | 128K | 1 | tools · json · vision · open-weights |
| Mercury 2 | $0.313 | $0.938 | 128K | 1 | tools · json · reasoning |
| Venice Role Play Uncensored | $0.500 | $2.00 | 128K | 1 | tools · json · vision · open-weights |
| Aion 3.0 Mini | $0.875 | $1.75 | 128K | 1 | tools · json · reasoning |
| Seed 2.1 Turbo | $0.625 | $3.13 | 256K | 1 | tools · json · reasoning · vision |
| Inkling | $1.25 | $5.06 | 524K | 1 | tools · json · reasoning · vision · open-weights |
| Aion 3.0 | $3.75 | $7.50 | 128K | 1 | tools · json · reasoning |
| GLM 4.7 Flash Hereticderivative | $0.070 | $0.400 | 200K | 1 | tools · json · reasoning · open-weights |
| Gemma 4 Uncensoredderivative | $0.163 | $0.500 | 256K | 1 | tools · json · vision · open-weights |
| DeepSeek V4 Flash 0731 Fastderivative | $0.350 | $0.700 | 1M | 1 | tools · json · reasoning · open-weights |
| MiniMax M2.5derivative | $0.270 | $0.950 | 198K | 1 | tools · reasoning · open-weights |
| MiniMax M3 Previewderivative | $0.300 | $1.20 | 524K | 1 | tools · reasoning · vision · open-weights |
| Qwen 3 Coder 480B Turboderivative | $0.350 | $1.50 | 256K | 1 | tools · json · open-weights |
| GPT-5.6 Lunaderivative | $0.267 | $1.60 | 1M | 1 | tools · json · reasoning · vision |
| MiniMax M2.7derivative | $0.375 | $1.50 | 198K | 1 | tools · reasoning · open-weights |
| Qwen 3 Next 80bderivative | $0.350 | $1.90 | 256K | 1 | tools · json · open-weights |
| Qwen 3.7 Plusderivative | $0.500 | $2.00 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.4 Miniderivative | $0.938 | $5.63 | 400K | 1 | tools · json · reasoning · vision |
| GPT-5.6 Luna Proderivative | $1.25 | $7.50 | 1M | 1 | tools · json · reasoning · vision |
| Qwen 3.8 Maxderivative | $2.50 | $7.50 | 1M | 1 | tools · reasoning · vision |
| Qwen 3.8 2.4Tderivative | $2.50 | $7.50 | 262K | 1 | tools · reasoning · open-weights |
| Qwen 3.7 Maxderivative | $2.70 | $8.05 | 1M | 1 | tools · reasoning · vision |
| GPT-5.3 Codexderivative | $2.19 | $17.50 | 400K | 1 | tools · json · reasoning · vision |
| GPT-5.2derivative | $2.19 | $17.50 | 256K | 1 | tools · json · reasoning |
| GPT-5.2 Codexderivative | $2.19 | $17.50 | 256K | 1 | tools · json · reasoning · vision |
| GPT-5.6 Terraderivative | $3.13 | $18.75 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.6 Terra Proderivative | $3.13 | $18.75 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.4derivative | $3.13 | $18.80 | 1M | 1 | tools · json · reasoning · vision |
| Kimi K3 Fastderivative | $4.50 | $22.50 | 1M | 1 | tools · json · reasoning · vision · open-weights |
| GPT-5.5derivative | $6.25 | $37.50 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.6 Sol Proderivative | $6.25 | $37.50 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.6 Solderivative | $6.25 | $37.50 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.4 Proderivative | $37.50 | $225.00 | 1M | 1 | tools · json · reasoning · vision |
| GPT-5.5 Proderivative | $37.50 | $225.00 | 1M | 1 | tools · json · reasoning · vision |
Frequently asked questions
How many AI models does venice offer?
We track 7 canonical venice models plus 27 community fine-tunes / derivatives (excluded from the main table). The list is recomputed daily.
Which venice model is the cheapest?
Venice Uncensored 1.2 is currently the lowest-priced venice model, at $0.200 per 1M input tokens and $0.900 per 1M output tokens. For the full apples-to-apples list, see /pricing/cheapest-llm-api.
Which venice model has the largest context window?
Inkling leads at 524K tokens. This is the total of prompt + completion.
Which venice models support tool calling?
Multiple venice models support tool calling, with Venice Uncensored 1.2 being a popular pick. The capability column in the table above marks every model with venice tool-calling support.
Which venice models accept image input?
Venice Uncensored 1.2 accepts image input. Other vision-capable venice models are tagged 'vision' in the table above. See /capabilities/vision for a cross-vendor comparison.
What are the best alternatives to venice?
Depends on the use case. For raw cost savings, look at /pricing/cheapest-llm-api. For agent-oriented workloads, /best/best-ai-model-for-agents. For long-document workflows, /best/best-long-context-llm.
How fresh is this venice pricing data?
Daily. Our pipeline syncs every morning and rebuilds these pages on data change, so list-price moves and new model releases land within roughly 24 hours.
Explore more
Top venice models
- Venice Uncensored 1.2$0.20 in / $0.90 out
- Mercury 2$0.31 in / $0.94 out
- Venice Role Play Uncensored$0.50 in / $2.00 out
- Aion 3.0 Mini$0.88 in / $1.75 out
- Seed 2.1 Turbo$0.63 in / $3.13 out
Browse by use case
Browse by capability
Last updated:
Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.
Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.