AI 모델 인텔리전스

기능 · 2026-09-28

롱 컨텍스트를 지원하는 AI 모델

200K 토큰 이상의 컨텍스트 윈도우를 가진 모델 비교.

이게 뭔가요?

  • 롱 컨텍스트 LLM은 단일 프롬프트에 200K 토큰 이상을 받을 수 있습니다 — 책 한 권, 멀티파일 저장소, 수 시간 분량의 트랜스크립트 등.
  • 일부 모델은 1M, 2M 토큰 이상으로 확장됩니다.

왜 중요한가

  • 롱 컨텍스트는 RAG를 보완하거나 대체합니다 — 검색된 조각만이 아닌 전체 콘텐츠를 붙여넣을 수 있습니다.
  • 유효 리콜은 길이에 따라 저하될 수 있으며, 긴 프롬프트는 백만 토큰당 가격으로 비용이 높아집니다.
  • 일부 제공사는 200K 이상에 단계별 요금을 적용합니다 — 각 모델 상세 페이지를 확인하세요.

이 기능을 지원하는 모델 1027개

모델벤더입력 / 1M출력 / 1M컨텍스트제공자
Ling 3.0 Flash VLopenrouter$0.021$0.062262K1
Ling 3.0 Flashopenrouter$0.021$0.063262K1
Ling 3.0 Flashvercel$0.021$0.063256K1
DeepSeek V4 Flash 0731DeepSeek$0.035$0.0701.31M47
Qwen3.5 4BAlibaba (Qwen)$0.040$0.070262K3
Model Routerazure$0.140Unknown200K1
Model Routerazure-cognitive-services$0.140Unknown200K1
Qwen3.7 FlashAlibaba (Qwen)$0.030$0.1301M13
Laguna XS 2.1openrouter$0.060$0.120262K1
GLM Flash LatestZ.AI / Zhipu$0.045$0.1401.31M4
Qwen3.5 9BAlibaba (Qwen)$0.040$0.150262K23
Mercury 2.5inception$0.040$0.150260K4
Mercury 2.5 Previewinception$0.040$0.150260K1
Nemotron 3.5 Lightning 30B A3BNVIDIA$0.050$0.150262K6
Agnes 3.0 Flashnano-gpt$0.050$0.150524K1
Space Bunny Alphanano-gpt$0.050$0.1501M1
DeepSeek V4 Flash LatestDeepSeek$0.050$0.1601.31M4
Nemotron 3.5 Lightning 30B A3BNVIDIA$0.039$0.1801M6
Greg 1 Minicrof$0.070$0.150229K1
Mercury 2.5venice$0.050$0.187260K1
Ling 3.0 Flash VLnano-gpt$0.060$0.180262K1
InclusionAI Ling 3.0 Flash (DeepInfra)llmgateway-providers$0.060$0.180262K1
InclusionAI Ling 3.0 Flash (NovitaAI)novita$0.060$0.180262K1
inclusionAI: Ling 3.0 Flashkilo$0.060$0.180262K1
inclusionAI: Ling 3.0 Flash Finkilo$0.060$0.180262K1
Ling 3.0 Flash Finopenrouter$0.060$0.180262K1
Ling 3.0 Flash Finvercel$0.060$0.180256K1
InclusionAI Ling 3.0 Flashllmgateway$0.060$0.180262K1
ministral-3b-2512cortecs$0.123$0.123256K1
Qwen TurboAlibaba (Qwen)$0.050$0.2001M6
Solar Mini 4Upstage$0.050$0.200524K3
Gemma 4 26B A4B ITGoogle$0.042$0.220262K22
Gemini-2.0-Flash-LiteGoogle$0.052$0.210990K3
Laguna S 2.1openrouter$0.090$0.1801.05M1
Ling 3.0 Flashnano-gpt$0.075$0.220262K1
Ling 3.0 Flash Thinkingnano-gpt$0.075$0.220262K1
inclusionAI: Ling 3.0 Flash VLkilo$0.075$0.220262K1
Ling 3.0 Flash VLvercel$0.075$0.220256K1
Amazon Nova Lite 1.0nano-gpt$0.059$0.238300K1
Ministral 3 8B 2512Mistral$0.150$0.150262K4
Amazon: Nova Lite 1.0kilo$0.060$0.240300K1
Nova Lite 1.0openrouter$0.060$0.240300K1
Nova Litevercel$0.060$0.240300K1
Nova Lite (US)edenai$0.060$0.240300K1
Nova Liteedenai$0.060$0.240300K1
Nova Lite (US)amazon-bedrock$0.060$0.240300K1
Nova Liteamazon-bedrock$0.060$0.240300K1
Ministral 8Bllmgateway$0.150$0.150262K1
Muse Spark 1.2 ContributorMeta$0.100$0.2001.05M6
Muse Spark 1.3 ContributorMeta$0.100$0.2001.05M5
Laguna S 2.1 Thinkingnano-gpt$0.100$0.2001.05M1
Laguna S 2.1nano-gpt$0.100$0.2001.05M1
Muse Spark 1.3 Contributorbothub$0.100$0.2001.05M1
Muse Spark 1.2 Contributor (Meta Contributor)llmgateway-providers$0.100$0.2001.05M1
Muse Spark 1.3 Contributor (Meta Contributor)llmgateway-providers$0.100$0.2001.05M1
Poolside: Laguna S 2.1kilo$0.100$0.2001.05M1
Poolside: Laguna XS 2.1kilo$0.100$0.200262K1
Laguna S 2.1vercel$0.100$0.2001M1
Laguna S 2.1pioneer$0.100$0.2001M1
Muse Spark 1.2 Contributoropencode-go$0.100$0.2001.05M1

전체 1027개 중 상위 60개 표시. 추가 필터링은 전체 목록을 이용하세요.

Frequently asked questions

How many AI models support 200K+ 컨텍스트?

1027 canonical models in our database currently support 200K+ 컨텍스트. The list is regenerated on every data refresh, so it always reflects the latest releases tracked in our catalogue.

What is the cheapest model with 200K+ 컨텍스트?

Ling 3.0 Flash VL from openrouter is currently the lowest-priced option, at $0.021 per 1M input tokens and $0.062 per 1M output tokens. The full table above is sorted price-ascending.

Which model with 200K+ 컨텍스트 has the largest context window?

Qwen Long (Alibaba (Qwen)) leads on context at 10M tokens. This may matter if you also need long-document understanding alongside 200K+ 컨텍스트.

Which models are available on the most providers?

Production-readiness usually correlates with how many independent providers host the same weights. The top three by provider count are: GLM-5.2 (89), Kimi K3 (75), DeepSeek V4 Pro (72).

How is 200K+ 컨텍스트 different from a regular LLM?

Long-context models accept ≥ 200K input tokens — enough for entire books, codebases or hours of transcripts in one prompt. Effective recall and per-token pricing both degrade with input length, so 'big context' is not always the right choice over RAG.

How often is this list updated?

Daily. Our data pipeline syncs once a day, regenerates the canonical model list, and rebuilds these pages so newly released models appear within 24 hours.

마지막 업데이트:

Prices in USD per 1M tokens. Unknown means the provider does not publish per-token pricing.

Pricing and capabilities are refreshed daily and reconciled against each provider's official documentation. Always verify critical production decisions with the provider directly.