Catalog data · Updated July 24, 2026

Vision LLM Models

Find active text-generation model identities that accept image input. Compare providers, context windows, known token prices, and related capabilities.

Objective catalog fields No overall quality score
335
Vision model identities
25
Providers shown
1
Data sections
6,154
Catalog records

Vision models with image input

Active text-generation model identities that accept image input.

Showing 50 of 335

Model Provider offers Why listed Context Input / output price Updated
Grok 4.3
grok-4-3
Venice
1 offer
Image input and text output 1,000,000
$1.42 / $2.83
per 1M tokens, when known
June 11, 2026
Grok 4.5
grok-4-5
Venice
1 offer
Image input and text output 500,000
$2.27 / $6.80
per 1M tokens, when known
July 8, 2026
Grok Build 0.1
grok-build-0-1
Venice
1 offer
Image input and text output 256,000
$1.00 / $2.00
per 1M tokens, when known
June 11, 2026
Grok Build 0.1
grok-build-0.1
Opencode, Openrouter, Xai +1
4 offers
Image input and text output 256,000
$1.00 / $2.00
per 1M tokens, when known
May 20, 2026
Inkling
inkling
Venice
1 offer
Image input and text output 1,000,000
$1.25 / $5.06
per 1M tokens, when known
July 17, 2026
Inkling
thinkingmachines/Inkling
Togetherai
1 offer
Image input and text output 524,288
$1.00 / $4.05
per 1M tokens, when known
July 15, 2026
Kimi K2.5
Kimi-K2.5
Togetherai
1 offer
Image input and text output 262,144
$0.50 / $2.80
per 1M tokens, when known
January 27, 2026
Kimi K2.5
kimi-k2-5
Venice
1 offer
Image input and text output 256,000
$0.56 / $3.50
per 1M tokens, when known
June 11, 2026
Kimi K2.5
kimi-k2.5
Moonshotai cn, Ollama cloud, Opencode +2
5 offers
Image input and text output 262,144
$0.57 / $2.85
per 1M tokens, when known
January 27, 2026
Kimi K2.6
Kimi-K2.6
Togetherai
1 offer
Image input and text output 262,144
$1.20 / $4.50
per 1M tokens, when known
April 21, 2026
Kimi K2.6
accounts/fireworks/models/kimi-k2p6
Fireworks ai
1 offer
Image input and text output 262,000
$0.95 / $4.00
per 1M tokens, when known
April 17, 2026
Kimi K2.6
kimi-k2-6
Venice
1 offer
Image input and text output 256,000
$0.75 / $3.50
per 1M tokens, when known
June 11, 2026
Kimi K2.6
kimi-k2.6
Cloudflare workers ai, Moonshotai, Moonshotai cn +4
7 offers
Image input and text output 262,144
$0.65 / $2.72
per 1M tokens, when known
April 21, 2026
Kimi K2.6 Fast
accounts/fireworks/routers/kimi-k2p6-fast
Fireworks ai
1 offer
Image input and text output 262,000
$2.00 / $8.00
per 1M tokens, when known
June 5, 2026
Kimi K2.6 Turbo
accounts/fireworks/routers/kimi-k2p6-turbo
Fireworks ai
1 offer
Image input and text output 262,000
$2.00 / $8.00
per 1M tokens, when known
April 17, 2026
Kimi K2.7 Code
accounts/fireworks/models/kimi-k2p7-code
Fireworks ai
1 offer
Image input and text output 262,000
$0.95 / $4.00
per 1M tokens, when known
June 16, 2026
Kimi K2.7 Code
kimi-k2-7-code
Venice
1 offer
Image input and text output 256,000
$0.75 / $3.50
per 1M tokens, when known
June 16, 2026
Kimi K2.7 Code
kimi-k2.7-code
Cloudflare workers ai, Moonshotai, Moonshotai cn +4
7 offers
Image input and text output 262,144
$0.78 / $3.50
per 1M tokens, when known
June 12, 2026
Kimi K2.7 Code (Free)
kimi-k2.7-code-free
Zenmux
1 offer
Image input and text output 262,144
$0.00 / $0.00
per 1M tokens, when known
June 12, 2026
Kimi K2.7 Code Fast
accounts/fireworks/routers/kimi-k2p7-code-fast
Fireworks ai
1 offer
Image input and text output 262,000
$1.90 / $8.00
per 1M tokens, when known
June 16, 2026
Kimi K2.7 Code HighSpeed
kimi-k2.7-code-highspeed
Moonshotai, Moonshotai cn
2 offers
Image input and text output 262,144
$1.90 / $8.00
per 1M tokens, when known
June 12, 2026
Kimi K3
kimi-k3
Moonshotai, Openrouter, Venice +1
4 offers
Image input and text output 1,048,576
$3.00 / $15.00
per 1M tokens, when known
July 17, 2026
Kimi K3 (Free)
kimi-k3-free
Zenmux
1 offer
Image input and text output 1,048,576
$0.00 / $0.00
per 1M tokens, when known
July 16, 2026
Llama 4 Maverick 17B 128E Instruct FP8
llama-4-maverick-17b-128e-instruct-fp8
Github models
1 offer
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
January 31, 2025
Llama 4 Scout 17B 16E
meta-llama/llama-4-scout-17b-16e-instruct
Groq
1 offer
Image input and text output 131,072
$0.11 / $0.34
per 1M tokens, when known
April 5, 2025
Llama 4 Scout 17B 16E Instruct
llama-4-scout-17b-16e-instruct
Cloudflare workers ai, Github models
2 offers
Image input and text output 131,000
$0.00 / $0.00
per 1M tokens, when known
April 5, 2025
Llama-3.2-11B-Vision-Instruct
llama-3.2-11b-vision-instruct
Cloudflare workers ai, Github models
2 offers
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
September 25, 2024
Llama-3.2-90B-Vision-Instruct
llama-3.2-90b-vision-instruct
Github models
1 offer
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
September 25, 2024
Meta: Llama 4 Maverick
llama-4-maverick
Openrouter
1 offer
Image input and text output 1,048,576
$0.20 / $0.80
per 1M tokens, when known
April 5, 2025
Meta: Llama 4 Scout
llama-4-scout
Openrouter
1 offer
Image input and text output 1,310,720
$0.10 / $0.30
per 1M tokens, when known
April 5, 2025
Meta: Llama Guard 4 12B
llama-guard-4-12b
Openrouter
1 offer
Image input and text output 1,048,576
$0.18 / $0.18
per 1M tokens, when known
April 30, 2025
Meta: Muse Spark 1.1
muse-spark-1.1
Openrouter
1 offer
Image input and text output 1,048,576
$1.25 / $4.25
per 1M tokens, when known
July 9, 2026
MiMo V2 Omni
mimo-v2-omni
Zenmux
1 offer
Image input and text output 265,000
$0.40 / $2.00
per 1M tokens, when known
March 18, 2026
MiMo V2.5 Free
mimo-v2.5-free
Opencode
1 offer
Image input and text output 200,000
$0.00 / $0.00
per 1M tokens, when known
April 24, 2026
MiMo-V2.5
xiaomi-mimo-v2-5
Venice
1 offer
Image input and text output 1,000,000
$0.14 / $0.28
per 1M tokens, when known
June 11, 2026
MiniMax M3 Preview
minimax-m3-preview
Venice
1 offer
Image input and text output 524,288
$0.30 / $1.20
per 1M tokens, when known
June 13, 2026
MiniMax-M3
MiniMax-M3
Minimax, Togetherai
2 offers
Image input and text output 1,000,000
$0.30 / $1.20
per 1M tokens, when known
June 25, 2026
MiniMax-M3
minimax-m3
Fireworks ai, Ollama cloud, Opencode +2
5 offers
Image input and text output 1,048,576
$0.30 / $1.20
per 1M tokens, when known
June 12, 2026
MiniMax: MiniMax-01
minimax-01
Openrouter
1 offer
Image input and text output 1,000,192
$0.20 / $1.10
per 1M tokens, when known
January 15, 2025
Mistral Large (latest)
mistral-large-latest
Mistral
1 offer
Image input and text output 262,144
$0.50 / $1.50
per 1M tokens, when known
December 2, 2025
Mistral Large 3
mistral-large-2512
Mistral
1 offer
Image input and text output 262,144
$0.50 / $1.50
per 1M tokens, when known
December 2, 2025
Mistral Medium (latest)
mistral-medium-latest
Mistral
1 offer
Image input and text output 262,144
$1.50 / $7.50
per 1M tokens, when known
April 29, 2026
Mistral Medium 3
mistral-medium-2505
Mistral
1 offer
Image input and text output 131,072
$0.40 / $2.00
per 1M tokens, when known
May 7, 2025
Mistral Medium 3 (25.05)
mistral-ai/mistral-medium-2505
Github models
1 offer
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
May 1, 2025
Mistral Medium 3.1
mistral-medium-2508
Mistral
1 offer
Image input and text output 262,144
$0.40 / $2.00
per 1M tokens, when known
August 12, 2025
Mistral Medium 3.5
mistral-medium-2604
Mistral
1 offer
Image input and text output 262,144
$1.50 / $7.50
per 1M tokens, when known
April 29, 2026
Mistral Small (latest)
mistral-small-latest
Mistral
1 offer
Image input and text output 256,000
$0.15 / $0.60
per 1M tokens, when known
March 16, 2026
Mistral Small 3.1
mistral-ai/mistral-small-2503
Github models
1 offer
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
March 1, 2025
Mistral Small 3.2
mistral-small-2506
Mistral
1 offer
Image input and text output 128,000
$0.10 / $0.30
per 1M tokens, when known
June 20, 2025
Mistral Small 4
mistral-small-2603
Mistral, Venice
2 offers
Image input and text output 256,000
$0.15 / $0.60
per 1M tokens, when known
June 11, 2026

How the vision list works

The page groups active text-generation offers that list image input and text output into conservative model identities.

Inclusion criteria

  • The provider offer has explicit typed text-generation support.
  • The input modality set includes image and the output modality set includes text.

Exclusions

  • The list excludes image-generation-only records that do not return text.
  • The list excludes catalog-only, deprecated, retired, and disallowed offers.

Data rules

  • Vision meaning: Vision means that the catalog lists image as an input modality. It does not mean image generation.
  • Grouping: Provider offers are grouped by the conservative model identity rule used by the LLM models directory.

Limits of the data

  • Image input metadata does not measure visual accuracy, supported image size, or document quality.
  • A missing modality can mean that the source does not have a confirmed value.

What vision means here

This page uses the catalog modality fields. A model is present when an active text offer accepts images and returns text. The list does not claim that every image model has the same level of visual reasoning.

Sources

  • llm_db — Source for typed execution, modality, provider, context, and price data. Checked 2026-07-30.