Catalog data · Updated July 24, 2026

Vision LLM Models

Find active text-generation model identities that accept image input. Compare providers, context windows, known token prices, and related capabilities.

Objective catalog fields No overall quality score
335
Vision model identities
25
Providers shown
1
Data sections
6,154
Catalog records

Vision models with image input

Active text-generation model identities that accept image input.

Showing 50 of 335

Model Provider offers Why listed Context Input / output price Updated
Google: Gemma 3 4B
google/gemma-3-4b-it
Openrouter
1 offer
Image input and text output 131,072
$0.05 / $0.10
per 1M tokens, when known
March 13, 2025
Google: Gemma 4 26B A4B
google/gemma-4-26b-a4b-it
Openrouter
1 offer
Image input and text output 262,144
$0.12 / $0.35
per 1M tokens, when known
April 2, 2026
Google: Gemma 4 26B A4B (free)
google/gemma-4-26b-a4b-it:free
Openrouter
1 offer
Image input and text output 262,144
$0.00 / $0.00
per 1M tokens, when known
April 2, 2026
Google: Gemma 4 31B
google/gemma-4-31b-it
Openrouter
1 offer
Image input and text output 262,144
$0.14 / $0.40
per 1M tokens, when known
April 2, 2026
Google: Gemma 4 31B (free)
google/gemma-4-31b-it:free
Openrouter
1 offer
Image input and text output 262,144
$0.00 / $0.00
per 1M tokens, when known
April 2, 2026
Google: Lyria 3 Clip Preview
lyria-3-clip-preview
Google, Openrouter
2 offers
Image input and text output 1,048,576
$0.00 / $0.00
per 1M tokens, when known
March 25, 2026
Google: Lyria 3 Pro Preview
lyria-3-pro-preview
Google, Openrouter
2 offers
Image input and text output 1,048,576
$0.00 / $0.00
per 1M tokens, when known
March 25, 2026
GPT-4.1
gpt-4.1
Github models, Nearai, Openai +1
4 offers
Image input and text output 1,047,576
$0.00 / $0.00
per 1M tokens, when known
April 14, 2025
GPT-4.1 mini
gpt-4.1-mini
Github models, Nearai, Openai +1
4 offers
Image input and text output 1,047,576
$0.00 / $0.00
per 1M tokens, when known
April 14, 2025
GPT-4o
gpt-4o
Github models, Openai, Openrouter +1
4 offers
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
August 6, 2024
GPT-4o
openai-gpt-4o-2024-11-20
Venice
1 offer
Image input and text output 128,000
$3.13 / $12.50
per 1M tokens, when known
June 11, 2026
GPT-4o mini
gpt-4o-mini
Github models, Openai, Openrouter +1
4 offers
Image input and text output 128,000
$0.00 / $0.00
per 1M tokens, when known
July 18, 2024
GPT-4o Mini
openai-gpt-4o-mini-2024-07-18
Venice
1 offer
Image input and text output 128,000
$0.19 / $0.75
per 1M tokens, when known
June 11, 2026
GPT-5
gpt-5
Openai, Opencode
2 offers
Image input and text output 400,000
$1.07 / $8.50
per 1M tokens, when known
August 7, 2025
GPT-5 Mini
gpt-5-mini
Nearai, Openai, Openrouter
3 offers
Image input and text output 400,000
$0.25 / $2.00
per 1M tokens, when known
August 7, 2025
GPT-5 Nano
gpt-5-nano
Nearai, Openai, Opencode +1
4 offers
Image input and text output 400,000
$0.05 / $0.40
per 1M tokens, when known
August 7, 2025
GPT-5.1
gpt-5.1
Nearai, Openai, Opencode +2
5 offers
Image input and text output 400,000
$1.07 / $8.50
per 1M tokens, when known
November 13, 2025
GPT-5.2
gpt-5.2
Nearai, Openai, Opencode +2
5 offers
Image input and text output 400,000
$1.75 / $14.00
per 1M tokens, when known
December 11, 2025
GPT-5.2 Codex
openai-gpt-52-codex
Venice
1 offer
Image input and text output 256,000
$2.19 / $17.50
per 1M tokens, when known
June 11, 2026
GPT-5.3 Codex
gpt-5.3-codex
Openai, Opencode, Openrouter
3 offers
Image input and text output 400,000
$1.75 / $14.00
per 1M tokens, when known
February 24, 2026
GPT-5.3 Codex
openai-gpt-53-codex
Venice
1 offer
Image input and text output 400,000
$2.19 / $17.50
per 1M tokens, when known
June 11, 2026
GPT-5.3 Codex Spark
gpt-5.3-codex-spark
Openai
1 offer
Image input and text output 128,000
$1.75 / $14.00
per 1M tokens, when known
February 5, 2026
GPT-5.4
gpt-5.4
Nearai, Openai, Opencode +2
5 offers
Image input and text output 1,050,000
$2.50 / $15.00
per 1M tokens, when known
March 20, 2026
GPT-5.4
openai-gpt-54
Venice
1 offer
Image input and text output 1,000,000
$3.13 / $18.80
per 1M tokens, when known
June 11, 2026
GPT-5.4 mini
gpt-5.4-mini
Nearai, Openai, Opencode +1
4 offers
Image input and text output 400,000
$0.75 / $4.50
per 1M tokens, when known
March 17, 2026
GPT-5.4 Mini
openai-gpt-54-mini
Venice
1 offer
Image input and text output 400,000
$0.94 / $5.63
per 1M tokens, when known
June 11, 2026
GPT-5.4 nano
gpt-5.4-nano
Nearai, Openai, Opencode +1
4 offers
Image input and text output 400,000
$0.20 / $1.25
per 1M tokens, when known
March 17, 2026
GPT-5.4 Pro
gpt-5.4-pro
Openai, Opencode, Openrouter +1
4 offers
Image input and text output 1,050,000
$30.00 / $180.00
per 1M tokens, when known
March 20, 2026
GPT-5.4 Pro
openai-gpt-54-pro
Venice
1 offer
Image input and text output 1,000,000
$37.50 / $225.00
per 1M tokens, when known
June 11, 2026
GPT-5.5
gpt-5.5
Nearai, Openai, Opencode +2
5 offers
Image input and text output 1,050,000
$5.00 / $30.00
per 1M tokens, when known
April 23, 2026
GPT-5.5
openai-gpt-55
Venice
1 offer
Image input and text output 1,000,000
$6.25 / $37.50
per 1M tokens, when known
June 11, 2026
GPT-5.5 Instant
gpt-5.5-instant
Zenmux
1 offer
Image input and text output 400,000
$5.00 / $30.00
per 1M tokens, when known
May 28, 2026
GPT-5.5 Pro
gpt-5.5-pro
Openai, Opencode, Openrouter +1
4 offers
Image input and text output 1,050,000
$30.00 / $180.00
per 1M tokens, when known
April 24, 2026
GPT-5.5 Pro
openai-gpt-55-pro
Venice
1 offer
Image input and text output 1,000,000
$37.50 / $225.00
per 1M tokens, when known
June 11, 2026
GPT-5.6
gpt-5.6
Openai
1 offer
Image input and text output 1,050,000
$5.00 / $30.00
per 1M tokens, when known
July 9, 2026
GPT-5.6 Luna
openai-gpt-56-luna
Venice
1 offer
Image input and text output 1,000,000
$1.25 / $7.50
per 1M tokens, when known
July 9, 2026
GPT-5.6 Luna Pro
openai-gpt-56-luna-pro
Venice
1 offer
Image input and text output 1,000,000
$1.25 / $7.50
per 1M tokens, when known
July 9, 2026
GPT-5.6 Sol
openai-gpt-56-sol
Venice
1 offer
Image input and text output 1,000,000
$6.25 / $37.50
per 1M tokens, when known
July 9, 2026
GPT-5.6 Sol Pro
openai-gpt-56-sol-pro
Venice
1 offer
Image input and text output 1,000,000
$6.25 / $37.50
per 1M tokens, when known
July 9, 2026
GPT-5.6 Terra
openai-gpt-56-terra
Venice
1 offer
Image input and text output 1,000,000
$3.13 / $18.75
per 1M tokens, when known
July 9, 2026
GPT-5.6 Terra Pro
openai-gpt-56-terra-pro
Venice
1 offer
Image input and text output 1,000,000
$3.13 / $18.75
per 1M tokens, when known
July 9, 2026
Grok 4
grok-4
Zenmux
1 offer
Image input and text output 256,000
$3.00 / $15.00
per 1M tokens, when known
July 9, 2025
Grok 4 Fast
grok-4-fast
Zenmux
1 offer
Image input and text output 2,000,000
$0.20 / $0.50
per 1M tokens, when known
September 19, 2025
Grok 4.1 Fast
grok-4.1-fast
Zenmux
1 offer
Image input and text output 2,000,000
$0.20 / $0.50
per 1M tokens, when known
November 20, 2025
Grok 4.1 Fast Non Reasoning
grok-4.1-fast-non-reasoning
Zenmux
1 offer
Image input and text output 2,000,000
$0.20 / $0.50
per 1M tokens, when known
November 20, 2025
Grok 4.20
grok-4-20
Venice
1 offer
Image input and text output 2,000,000
$1.42 / $2.83
per 1M tokens, when known
June 11, 2026
Grok 4.20 (Non-Reasoning)
grok-4.20-0309-non-reasoning
Xai
1 offer
Image input and text output 1,000,000
$1.25 / $2.50
per 1M tokens, when known
March 9, 2026
Grok 4.20 (Reasoning)
grok-4.20-0309-reasoning
Xai
1 offer
Image input and text output 1,000,000
$1.25 / $2.50
per 1M tokens, when known
March 9, 2026
Grok 4.20 Multi-Agent
grok-4-20-multi-agent
Venice
1 offer
Image input and text output 2,000,000
$1.42 / $2.83
per 1M tokens, when known
June 11, 2026
Grok 4.20 Multi-Agent
grok-4.20-multi-agent-0309
Xai
1 offer
Image input and text output 1,000,000
$1.25 / $2.50
per 1M tokens, when known
March 9, 2026

How the vision list works

The page groups active text-generation offers that list image input and text output into conservative model identities.

Inclusion criteria

  • The provider offer has explicit typed text-generation support.
  • The input modality set includes image and the output modality set includes text.

Exclusions

  • The list excludes image-generation-only records that do not return text.
  • The list excludes catalog-only, deprecated, retired, and disallowed offers.

Data rules

  • Vision meaning: Vision means that the catalog lists image as an input modality. It does not mean image generation.
  • Grouping: Provider offers are grouped by the conservative model identity rule used by the LLM models directory.

Limits of the data

  • Image input metadata does not measure visual accuracy, supported image size, or document quality.
  • A missing modality can mean that the source does not have a confirmed value.

What vision means here

This page uses the catalog modality fields. A model is present when an active text offer accepts images and returns text. The list does not claim that every image model has the same level of visual reasoning.

Sources

  • llm_db — Source for typed execution, modality, provider, context, and price data. Checked 2026-07-30.