AI 模型情報

價格 · 2026-08-15

最便宜的 LLM API

按每 token 總價從低到高排序的文字 LLM API 列表。

關於這份榜單

  • 所有「文字進 / 文字出」的 LLM API,按每 token 綜合成本從便宜到貴排序。
  • 排除 $0 佔位定價(免費推廣檔、GitHub Copilot 重發等)——「未公開」不等於免費。
  • 用本列表在仍滿足上下文與能力需求的前提下,找到成本最低的選項。
#模型廠商輸入 / 1M輸出 / 1M總計上下文
1Voxtral Small 24B 2507Mistral$0.002$0.002$0.00532K
2All-MiniLM-L6-v2digitalocean$0.009Unknown$0.009256
3Multi-QA-mpnet-base-dot-v1digitalocean$0.009Unknown$0.009512
4Qwen3 Embedding 8BAlibaba (Qwen)$0.010Unknown$0.01033K
5Qwen3 Embedding 4BAlibaba (Qwen)$0.010Unknown$0.01033K
6Qwen3 Embedding 0.6BAlibaba (Qwen)$0.010Unknown$0.01033K
7BGE Reranker v2 M3digitalocean$0.010Unknown$0.0108K
8Llama 3.2 1B InstructMeta$0.010$0.010$0.02060K
9text-embedding-3-smallOpenAI$0.020Unknown$0.0208K
10Prompt Guard 2 86MMeta$0.010$0.010$0.020512
11Llama Prompt Guard 2 22MMeta$0.010$0.010$0.020512
12BGE M3digitalocean$0.020Unknown$0.0208K
13E5 Large v2digitalocean$0.020Unknown$0.020512
14text-embedding-3-smallsap-ai-core$0.020Unknown$0.0208K
15text-embedding-3-smallazure-cognitive-services$0.020Unknown$0.0208K
16text-embedding-3-smallazure$0.020Unknown$0.0208K
17Llama 3.2 3B InstructMeta$0.020$0.020$0.040131K
18Ling-2.6-flashopenrouter$0.010$0.030$0.040262K
19PaddleOCR-VLnovita-ai$0.020$0.020$0.04016K
20Llama-3.1-8B-InstructMeta$0.020$0.030$0.050131K
21Mistral Nemo Instruct 2407Mistral$0.020$0.030$0.050131K
22Nomic Embed Text v1.5tinfoil$0.050Unknown$0.0508K
23Meta Llama 3.1 8B Instruct TurboMeta$0.020$0.030$0.050128K
24DeepSeek OCR 2DeepSeek$0.030$0.030$0.0604K
25nvidia--llama-3.2-nv-embedqa-1bMeta$0.070Unknown$0.0708K
26Ministral 3Bazure-cognitive-services$0.040$0.040$0.080128K
27Ministral 3Bazure$0.040$0.040$0.080128K
28Ministral 3B (latest)Mistral$0.040$0.040$0.080128K
29Llama 3 8B InstructMeta$0.040$0.040$0.0808K
30Ling-3.0-flashopenrouter$0.021$0.063$0.084262K
31Llama 3 8B LunarisMeta$0.040$0.050$0.0908K
32GTE Large (v1.5)digitalocean$0.090Unknown$0.0908K
33text-embedding-3-largesap-ai-core$0.090Unknown$0.0908K
34text-embedding-ada-002OpenAI$0.100Unknown$0.1008K
35Mistral EmbedMistral$0.100Unknown$0.1008K
36text-embedding-ada-002azure-cognitive-services$0.100Unknown$0.1008K
37text-embedding-ada-002azure$0.100Unknown$0.1008K
38Sao10k L3 8B Lunaris novita-ai$0.050$0.050$0.1008K
39L3 8B Stheno V3.2novita-ai$0.050$0.050$0.1008K
40Llama 3.2 11B Vision InstructMeta$0.055$0.055$0.110128K
41Llama Guard 3 8BMeta$0.055$0.055$0.110131K
42Qwen3.5 4BAlibaba (Qwen)$0.040$0.070$0.11033K
43Gemma 3 4B ITGoogle$0.040$0.080$0.120128K
44MythoMax 13Bopenrouter$0.060$0.060$0.1208K
45MythoMax 13Bkilo$0.060$0.060$0.1204K
46Gemma 4 E4B ITGoogle$0.020$0.100$0.120131K
47Sarvam 30Bfastrouter$0.020$0.100$0.120128K
48Nex N2 Mininano-gpt$0.025$0.100$0.125262K
49Nex-N2-Miniopenrouter$0.025$0.100$0.125262K
50Nex AGI: Nex-N2-Minikilo$0.025$0.100$0.125262K
51Granite 4.0 H Microcloudflare-workers-ai$0.017$0.112$0.129131K
52Granite 4.0 Microopenrouter$0.017$0.112$0.129131K
53Granite 4.0 H Microcloudflare-ai-gateway$0.017$0.112$0.129131K
54IBM: Granite 4.0 Microkilo$0.017$0.112$0.129131K
55text-embedding-3-largeOpenAI$0.130Unknown$0.1308K
56Llama 3.1 8BMeta$0.050$0.080$0.130131K
57amazon--nova-microsap-ai-core$0.030$0.100$0.130128K
58text-embedding-3-largeazure-cognitive-services$0.130Unknown$0.1308K
59text-embedding-3-largeazure$0.130Unknown$0.1308K
60Sarvam 30Bnano-gpt$0.028$0.111$0.13966K
61amazon--titan-embed-textsap-ai-core$0.140Unknown$0.1408K
62Model Routerazure-cognitive-services$0.140Unknown$0.140200K
63Model Routerazure$0.140Unknown$0.140200K
64baichuan-m2-32bnovita-ai$0.070$0.070$0.140131K
65Qwen3.7 FlashAlibaba (Qwen)$0.030$0.118$0.1481M
66Google Gemma 3 12BGoogle$0.050$0.100$0.150131K
67Gemini Embedding 001Google$0.150Unknown$0.1502K
68LFM2-24B-A2Btogetherai$0.030$0.120$0.15033K
69LFM2 24B A2Bpioneer$0.030$0.120$0.15033K
70Granite 4.1 8Bnano-gpt$0.050$0.100$0.150131K
71Granite 4.1 8Bopenrouter$0.050$0.100$0.150131K
72IBM: Granite 4.1 8Bkilo$0.050$0.100$0.150131K
73Mellum2 12B A2.5Bwandb$0.050$0.100$0.150131K
74Granite 4.1 8Bwandb$0.050$0.100$0.150131K
75gpt-oss-20bOpenAI$0.030$0.130$0.160128K
76DeepSeek R1 Distill Llama 70BMeta$0.030$0.130$0.16033K
77R1 Distill Llama 70BDeepSeek$0.030$0.140$0.1708K
78GPT OSS 120Bllmgateway$0.032$0.140$0.172131K
79Nova Micro 1.0openrouter$0.035$0.140$0.175128K
80Nova Microamazon-bedrock$0.035$0.140$0.175128K

顯示前 80 項,共 1150 項。 查看其餘請用 完整目錄

Frequently asked questions

What is the cheapest LLM API right now?

Voxtral Small 24B 2507 is the lowest-priced AI APIs on this list, at $0.002 per 1M input tokens and $0.002 per 1M output tokens. The total column above sums input + output per 1M tokens for direct comparison.

Why are prices shown per 1M tokens?

Almost every commercial LLM provider publishes rates per million tokens. Per-1K-tokens or per-token rates are easy to misread (off by 1000×). Standardising on $/1M lets you compare vendors directly.

Are these prices including or excluding tax?

All prices are the provider's headline list rate in USD, before any volume discounts, prepaid credits, batch-API discounts, prompt-caching discounts or jurisdictional sales tax. Always verify with the provider before committing to a budget.

How often do these prices change?

Vendor list-price moves are typically picked up within hours of an announcement, and our pipeline re-syncs daily. Each change is also written to /changelog so you can audit historical pricing over time.

Why are some models showing 'Unknown' instead of a price?

We deliberately do not coerce missing data to $0. 'Unknown' means the provider does not publish a public rate (often models behind enterprise sales or invite-only access). Treating Unknown as free would push paid-but-unpriced models to the top of every cheap list — broken UX and broken SEO.

最近更新:

Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.

Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.