Quick:
nano_gpt
Active
Gemini 2.5 Flash Lite
Model Spec
nano_gpt:gemini-2.5-flash-lite
Modalities
Input
text
image
Output
text
Capabilities
Chat
Reasoning
Vision
Streaming
Pricing Components
| Component | Price | Conditions | Source |
|---|---|---|---|
| token.input | $0.10 / 1M tokens | Standard | — |
| token.output | $0.40 / 1M tokens | Standard | — |
Specifications
Context Window
1,048,756 tokens
Max Output
65,536 tokens
Input Cost
$0.10/M
Output Cost
$0.40/M
History
Raw JSONTimeline is based on collected llm_db history snapshots and may be incomplete.
2026-08-04
changed
-
cost.cache_read
(add)
— → 0.01
-
extra.description
(replace)
Compact GPT model for low-latency assistance and high-volume workloads → Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
-
extra.family
(add)
— → gemini-flash-lite
2026-07-15
changed
-
extra.reasoning_options
(replace)
[2 items] → [1 items]
2026-07-02
changed
-
extra.description
(add)
— → Compact GPT model for low-latency assistance and high-volume workloads
2026-07-01
changed
-
extra.reasoning_options
(replace)
[3 items] → [2 items]
2026-06-28
changed
-
extra.reasoning_options
(add)
— → [3 items]
2026-04-20
changed
-
capabilities.rerank
(add)
— → false
2026-03-27
changed
-
catalog_only
(add)
— → true
2026-03-14
introduced
snapshots 64bf744 -> b7ca903 • generated 2026-08-05T03:31:59
6172
of 6306 models
1 / 124
Page 1 of 124