Fournisseur · 2026-08-14
baseten
| Modèle | Entrée / 1M | Sortie / 1M | Contexte | Fournisseurs | Tags |
|---|---|---|---|---|---|
| Inkling Small | $0.500 | $1.20 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
| Inkling | $1.00 | $4.05 | 1.05M | 1 | tools · json · reasoning · vision · open-weights |
| Nemotron Superdérivé | $0.300 | $0.750 | 203K | 1 | tools · json · reasoning · open-weights |
Frequently asked questions
How many AI models does baseten offer?
We track 2 canonical baseten models plus 1 community fine-tunes / derivatives (excluded from the main table). The list is recomputed daily.
Which baseten model is the cheapest?
Inkling Small is currently the lowest-priced baseten model, at $0.500 per 1M input tokens and $1.20 per 1M output tokens. For the full apples-to-apples list, see /pricing/cheapest-llm-api.
Which baseten model has the largest context window?
Inkling Small leads at 1.05M tokens. This is the total of prompt + completion.
Which baseten models support tool calling?
Multiple baseten models support tool calling, with Inkling Small being a popular pick. The capability column in the table above marks every model with baseten tool-calling support.
Which baseten models accept image input?
Inkling Small accepts image input. Other vision-capable baseten models are tagged 'vision' in the table above. See /capabilities/vision for a cross-vendor comparison.
What are the best alternatives to baseten?
Depends on the use case. For raw cost savings, look at /pricing/cheapest-llm-api. For agent-oriented workloads, /best/best-ai-model-for-agents. For long-document workflows, /best/best-long-context-llm.
How fresh is this baseten pricing data?
Daily. Our pipeline syncs every morning and rebuilds these pages on data change, so list-price moves and new model releases land within roughly 24 hours.
Explore more
Top baseten models
- Inkling Small$0.50 in / $1.20 out
- Inkling$1.00 in / $4.05 out
Browse by use case
Browse by capability
Dernière mise à jour :
Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.
Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.