AI 模型情報

工具 · 2026-09-28

LLM 計費器

依您的 token 用量估算各模型的月度費用。

Presets:

Total tokens / month: 465,000,000 (360,000,000 in + 105,000,000 out)

#ModelVendorIn $/1MOut $/1MMonthly costvs cheapest
1gpt-oss-20bOpenAI$0.018$0.090$15.93—
2DeepSeek V4 Flash 0731DeepSeek$0.035$0.070$19.951.3×
3gpt-oss-120bOpenAI$0.030$0.170$28.651.8×
4Llama-3.3-70B-InstructMeta$0.050$0.230$42.152.6×
5GPT-5 NanoOpenAI$0.050$0.400$60.003.8×
6Gemma 4 31B ITGoogle$0.100$0.250$62.253.9×
7GLM-4.7-FlashZ.AI / Zhipu$0.060$0.400$63.604.0×
8Qwen3.8 27BAlibaba (Qwen)$0.080$0.350$65.554.1×
9Gemini 2.5 Flash-LiteGoogle$0.100$0.400$78.004.9×
10DeepSeek V4.1 FlashDeepSeek$0.140$0.280$79.805.0×
11GPT-6 LunaOpenAI$0.100$0.500$88.505.6×
12DeepSeek-V3.2DeepSeek$0.180$0.350$101.556.4×
13Qwen3.8 FlashAlibaba (Qwen)$0.150$0.470$103.356.5×
14GLM-5.3-FlashZ.AI / Zhipu$0.150$0.500$106.506.7×
15DeepSeek V4 FlashDeepSeek$0.150$0.600$117.007.3×
16GPT-5.6 LunaOpenAI$0.200$1.20$198.0012.4×
17GPT-5.4 nanoOpenAI$0.200$1.25$203.2512.8×
18DeepSeek V4 Pro 0813DeepSeek$0.350$0.800$210.0013.2×
19MiniMax-M3MiniMax$0.300$1.20$234.0014.7×
20MiniMax-M2.7MiniMax$0.300$1.20$234.0014.7×
21MiniMax-M2.5MiniMax$0.300$1.20$234.0014.7×
22Qwen3.6 35B-A3BAlibaba (Qwen)$0.248$1.49$245.2115.4×
23Gemini 3.1 Flash LiteGoogle$0.250$1.50$247.5015.5×
24DeepSeek V4 ProDeepSeek$0.435$0.870$247.9515.6×
25GPT-5 MiniOpenAI$0.250$2.00$300.0018.8×
26Kimi K2.5Moonshot AI$0.300$1.90$307.5019.3×
27Qwen3.7 PlusAlibaba (Qwen)$0.400$1.60$312.0019.6×
28GPT-4.1 miniOpenAI$0.400$1.60$312.0019.6×
29Gemini 2.5 FlashGoogle$0.300$2.50$370.5023.3×
30Gemini 3.5 Flash LiteGoogle$0.300$2.50$370.5023.3×
31GLM-4.7Z.AI / Zhipu$0.600$2.20$447.0028.1×
32GLM-4.6Z.AI / Zhipu$0.600$2.20$447.0028.1×
33Qwen3.6 PlusAlibaba (Qwen)$0.500$3.00$495.0031.1×
34Gemini 3 Flash PreviewGoogle$0.500$3.00$495.0031.1×
35Qwen3.5 397B-A17BAlibaba (Qwen)$0.600$3.60$594.0037.3×
36Qwen3.6 27BAlibaba (Qwen)$0.600$3.60$594.0037.3×
37Gemini 3.6 FlashGoogle$0.750$3.75$663.7541.7×
38Gemini 3.7 FlashGoogle$0.750$3.75$663.7541.7×
39Gemini 3.8 FlashGoogle$0.750$3.75$663.7541.7×
40GLM-5Z.AI / Zhipu$1.00$3.20$696.0043.7×
41Grok 4.3xAI$1.25$2.50$712.5044.7×
42GPT-5.4 miniOpenAI$0.750$4.50$742.5046.6×
43Kimi K2.6Moonshot AI$0.950$4.00$762.0047.8×
44Kimi K2.7 CodeMoonshot AI$0.950$4.00$762.0047.8×
45Claude Haiku 4.5 (latest)Anthropic$1.00$5.00$885.0055.6×
46GLM-5.2Z.AI / Zhipu$1.40$4.40$966.0060.6×
47GLM-5.3Z.AI / Zhipu$1.40$4.40$966.0060.6×
48GLM-5.1Z.AI / Zhipu$1.40$4.40$966.0060.6×
49Qwen3 MaxAlibaba (Qwen)$1.20$6.00$1,06266.7×
50Qwen3.8 MaxAlibaba (Qwen)$2.00$6.00$1,35084.7×
51Grok 4.6xAI$2.00$6.00$1,35084.7×
52Grok 4.5xAI$2.00$6.00$1,35084.7×
53Gemini 3.5 FlashGoogle$1.50$9.00$1,48593.2×
54GPT-5OpenAI$1.25$10.00$1,50094.2×
55Gemini 2.5 ProGoogle$1.25$10.00$1,50094.2×
56GPT-5.1OpenAI$1.25$10.00$1,50094.2×
57GPT-4.1OpenAI$2.00$8.00$1,56097.9×
58Qwen3.7 MaxAlibaba (Qwen)$2.50$7.50$1,688105.9×
59Claude Sonnet 5Anthropic$2.00$10.00$1,770111.1×
60GPT-6 SolOpenAI$2.00$10.00$1,770111.1×
61GPT-4oOpenAI$2.50$10.00$1,950122.4×
62GPT-5.6 TerraOpenAI$2.00$12.00$1,980124.3×
63Gemini 3.1 Pro PreviewGoogle$2.00$12.00$1,980124.3×
64GPT-5.2OpenAI$1.75$14.00$2,100131.8×
65GPT-5.3 CodexOpenAI$1.75$14.00$2,100131.8×
66GPT-5.4OpenAI$2.50$15.00$2,475155.4×
67Kimi K3Moonshot AI$3.00$15.00$2,655166.7×
68Claude Sonnet 4.6Anthropic$3.00$15.00$2,655166.7×
69Claude Sonnet 4.5 (latest)Anthropic$3.00$15.00$2,655166.7×
70GPT-5.6 SolOpenAI$4.00$20.00$3,540222.2×
71Claude Opus 5.5Anthropic$4.00$20.00$3,540222.2×
72Claude Opus 4.8Anthropic$5.00$25.00$4,425277.8×
73Claude Opus 4.7Anthropic$5.00$25.00$4,425277.8×
74Claude Opus 4.6Anthropic$5.00$25.00$4,425277.8×
75Claude Opus 5Anthropic$5.00$25.00$4,425277.8×
76Claude Opus 4.5 (latest)Anthropic$5.00$25.00$4,425277.8×
77GPT-5.5OpenAI$5.00$30.00$4,950310.7×
78Claude Fable 5Anthropic$10.00$50.00$8,850555.6×
79Claude Fable 5.1Anthropic$10.00$50.00$8,850555.6×
80GPT-6 AstraOpenAI$10.00$50.00$8,850555.6×

計算方式說明

monthly_cost = requests × ((avg_input_tokens × input_price + avg_output_tokens × output_price) / 1,000,000).

Numbers shown are estimates only. Real-world cost depends on prompt caching, >200K context tier rates, audio/image surcharges and provider-specific overage. Always confirm with the model detail page or the provider's official pricing.

Frequently asked questions

How accurate are the cost estimates?

We multiply your average input/output token counts by the recommended provider's headline rate from our daily-refreshed canonical catalogue. The math is exact, but the result is only as accurate as your token-count assumptions. Real production cost will also depend on prompt caching, batch-API discounts, >200K-context tier rates and audio/image surcharges — none of which the calculator factors in.

Where do the per-token prices come from?

All prices reflect each provider's published list rate in USD, refreshed daily from our normalised catalogue. They are NOT including taxes, prepaid credits or volume discounts.

Can I share or bookmark a calculation?

Yes. The URL captures your selected model and token settings, so you can link a teammate to a specific scenario without retyping the inputs.

Why are some models missing from the picker?

We only include text-LLMs with a published per-token price. Models without published rates (often invite-only or enterprise-gated) and embedding/audio-only models are excluded — the calculator is text-completion-focused.

How do I estimate cost for caching or long context?

Open the model's detail page — the Pricing detail block lists cache_read / cache_write / context_over_200k rates where applicable. The headline calculator number is a 'no-cache, ≤200K' baseline.

Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.

Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.