Catalog data · Updated July 24, 2026

Largest Context Window LLMs

Compare active text-generation model identities with recorded context windows of at least 128,000 tokens. Large outlier values are excluded from the ranking.

Objective catalog fields No overall quality score
595
Long-context identities
26
Providers shown
1
Data sections
6,154
Catalog records

Largest context windows

Active text-generation identities from 128,000 through 10 million tokens.

Showing 50 of 595

Model Provider offers Why listed Context Input / output price Updated
ByteDance: UI-TARS 7B
bytedance/ui-tars-1.5-7b
Openrouter
1 offer
128000 token context 128,000
$0.10 / $0.20
per 1M tokens, when known
July 22, 2025
Cohere Command A
cohere-command-a
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
November 1, 2024
Cohere Command R
cohere-command-r
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
August 1, 2024
Cohere Command R 08-2024
cohere-command-r-08-2024
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
August 1, 2024
Cohere Command R+
cohere/cohere-command-r-plus
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
August 1, 2024
Cohere Command R+ 08-2024
cohere/cohere-command-r-plus-08-2024
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
August 1, 2024
Cohere: Command R (08-2024)
command-r-08-2024
Cohere, Openrouter
2 offers
128000 token context 128,000
$0.15 / $0.60
per 1M tokens, when known
August 30, 2024
Cohere: Command R+ (08-2024)
cohere/command-r-plus-08-2024
Openrouter
1 offer
128000 token context 128,000
$2.50 / $10.00
per 1M tokens, when known
August 30, 2024
Cohere: Command R7B (12-2024)
command-r7b-12-2024
Cohere, Openrouter
2 offers
128000 token context 128,000
$0.04 / $0.15
per 1M tokens, when known
December 2, 2024
Command A Plus
command-a-plus-05-2026
Cohere
1 offer
128000 token context 128,000
$2.50 / $10.00
per 1M tokens, when known
June 9, 2026
Command A Vision
command-a-vision-07-2025
Cohere
1 offer
128000 token context 128,000
$2.50 / $10.00
per 1M tokens, when known
July 31, 2025
Command R+
command-r-plus-08-2024
Cohere
1 offer
128000 token context 128,000
$2.50 / $10.00
per 1M tokens, when known
August 30, 2024
Command R7B Arabic
command-r7b-arabic-02-2025
Cohere
1 offer
128000 token context 128,000
$0.04 / $0.15
per 1M tokens, when known
February 27, 2025
Deep Cogito: Cogito v2.1 671B
cogito-v2.1-671b
Openrouter
1 offer
128000 token context 128,000
$1.25 / $1.25
per 1M tokens, when known
November 13, 2025
DeepSeek-V3-0324
deepseek-v3-0324
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
March 24, 2025
Devstral Medium
devstral-medium-2507
Mistral
1 offer
128000 token context 128,000
$0.40 / $2.00
per 1M tokens, when known
July 10, 2025
Devstral Small
devstral-small-2507
Mistral
1 offer
128000 token context 128,000
$0.10 / $0.30
per 1M tokens, when known
July 10, 2025
Devstral Small 2505
devstral-small-2505
Mistral
1 offer
128000 token context 128,000
$0.10 / $0.30
per 1M tokens, when known
May 7, 2025
Gemma Sea Lion V4 27B It
gemma-sea-lion-v4-27b-it
Cloudflare workers ai
1 offer
128000 token context 128,000
$0.35 / $0.56
per 1M tokens, when known
September 23, 2025
GLM 4.7 Flash
zai-org-glm-4.7-flash
Venice
1 offer
128000 token context 128,000
$0.13 / $0.50
per 1M tokens, when known
June 11, 2026
GPT-4o
gpt-4o
Github models, Openai, Openrouter
3 offers
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
August 6, 2024
GPT-4o
openai-gpt-4o-2024-11-20
Venice
1 offer
128000 token context 128,000
$3.13 / $12.50
per 1M tokens, when known
June 11, 2026
GPT-4o mini
gpt-4o-mini
Github models, Openai, Openrouter
3 offers
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
July 18, 2024
GPT-4o Mini
openai-gpt-4o-mini-2024-07-18
Venice
1 offer
128000 token context 128,000
$0.19 / $0.75
per 1M tokens, when known
June 11, 2026
GPT-5.3 Codex Spark
gpt-5.3-codex-spark
Openai, Opencode
2 offers
128000 token context 128,000
$1.75 / $14.00
per 1M tokens, when known
February 12, 2026
Grok 3
grok-3
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
December 9, 2024
Grok 3 Mini
grok-3-mini
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
December 9, 2024
Hermes 3 Llama 3.1 405b
hermes-3-llama-3.1-405b
Venice
1 offer
128000 token context 128,000
$1.10 / $3.00
per 1M tokens, when known
June 11, 2026
Inception: Mercury 2
mercury-2
Openrouter, Venice
2 offers
128000 token context 128,000
$0.25 / $0.75
per 1M tokens, when known
June 11, 2026
Ling-1T
ling-1t
Zenmux
1 offer
128000 token context 128,000
$0.56 / $2.24
per 1M tokens, when known
October 9, 2025
Llama 3.2 3B
llama-3.2-3b
Venice
1 offer
128000 token context 128,000
$0.15 / $0.60
per 1M tokens, when known
June 11, 2026
Llama 3.3 70B
llama-3.3-70b
Venice
1 offer
128000 token context 128,000
$0.70 / $2.80
per 1M tokens, when known
June 11, 2026
Llama 4 Maverick 17B 128E Instruct FP8
llama-4-maverick-17b-128e-instruct-fp8
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
January 31, 2025
Llama-3.2-11B-Vision-Instruct
llama-3.2-11b-vision-instruct
Cloudflare workers ai, Github models
2 offers
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
September 25, 2024
Llama-3.2-90B-Vision-Instruct
llama-3.2-90b-vision-instruct
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
September 25, 2024
Magistral Medium (latest)
magistral-medium-latest
Mistral
1 offer
128000 token context 128,000
$2.00 / $5.00
per 1M tokens, when known
March 20, 2025
Magistral Small
magistral-small
Mistral
1 offer
128000 token context 128,000
$0.50 / $1.50
per 1M tokens, when known
March 17, 2025
Meta-Llama-3.1-405B-Instruct
meta-llama-3.1-405b-instruct
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
July 23, 2024
Meta-Llama-3.1-70B-Instruct
meta-llama-3.1-70b-instruct
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
July 23, 2024
Meta-Llama-3.1-8B-Instruct
meta-llama-3.1-8b-instruct
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
July 23, 2024
Ministral 3B
ministral-3b
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
October 22, 2024
Ministral 3B (latest)
ministral-3b-latest
Mistral
1 offer
128000 token context 128,000
$0.04 / $0.04
per 1M tokens, when known
October 4, 2024
Ministral 8B (latest)
ministral-8b-latest
Mistral
1 offer
128000 token context 128,000
$0.10 / $0.10
per 1M tokens, when known
October 4, 2024
Mistral Large
mistral-large
Openrouter
1 offer
128000 token context 128,000
$2.00 / $6.00
per 1M tokens, when known
February 26, 2024
Mistral Large 24.11
mistral-ai/mistral-large-2411
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
November 1, 2024
Mistral Medium 3 (25.05)
mistral-ai/mistral-medium-2505
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
May 1, 2025
Mistral Small 3.1
mistral-ai/mistral-small-2503
Github models
1 offer
128000 token context 128,000
$0.00 / $0.00
per 1M tokens, when known
March 1, 2025
Mistral Small 3.1 24B Instruct
mistral-small-3.1-24b-instruct
Cloudflare workers ai
1 offer
128000 token context 128,000
$0.35 / $0.56
per 1M tokens, when known
March 18, 2025
Mistral Small 3.2
mistral-small-2506
Mistral
1 offer
128000 token context 128,000
$0.10 / $0.30
per 1M tokens, when known
June 20, 2025
Mistral: Mistral Small 3.1 24B
mistralai/mistral-small-3.1-24b-instruct
Openrouter
1 offer
128000 token context 128,000
$0.35 / $0.56
per 1M tokens, when known
March 17, 2025

How the context ranking works

The page orders active text-generation identities by their largest recorded provider context limit.

Inclusion criteria

  • The model identity has at least one offer that passes the active text-generation rules.
  • The largest recorded context window is from 128,000 through 10 million tokens.

Exclusions

  • The list excludes missing and nonnumeric context values.
  • The list excludes values above 10 million tokens as unverified outliers.

Data rules

  • Context value: A grouped identity uses the largest recorded context limit among its eligible provider offers.
  • Ordering: Model identities are ordered by context size, then name and model ID for stable output.

Limits of the data

  • A large context window does not prove strong recall, reasoning, speed, or low cost across the full window.
  • Providers can apply different context limits to the same model identity.

Read context limits with care

The advertised limit is only one part of long-context performance. Test retrieval quality, latency, and total prompt cost with data that matches your application.

Sources

  • llm_db — Source for active offers and recorded model context limits. Checked 2026-07-30.