Gemini 3.5 Flash Lite
google · google/gemini-3.5-flash-lite
Fast Gemini model balancing multimodal reasoning, tool use, and cost
≈699k Chinese characters of context · Reads images · Step-by-step reasoning · Calls tools / agents · Structured JSON output
Read a 200,000-Chinese-character book and write a 2,000-Chinese-character summary: about $0.10
An order-of-magnitude estimate at 1.5 tokens per Chinese character and the vendor's own rate — not a quote. Full method: the glossary on the AI Models page.
Specs
| Context | 1,048,576(≈699k Chinese characters) |
|---|---|
| Max output | 65,536 |
| Structured output | Yes |
| Open weights | No |
| Released | 2026-07-21 |
Source: models.dev snapshot 2026-08-29
Who serves it, at what price
| Provider | In /M | Out /M | Cache read /M |
|---|---|---|---|
| Googlefirst-party | $0.3 | $2.5 | $0.03 |
| Abacus | $0.3 | $2.5 | $0.03 |
| Cortecs | $0.33 | $2.749 | $0.033 |
| Vertex | $0.3 | $2.5 | $0.03 |
| DevPass (LLM Gateway) | $0.3 | $2.5 | $0.03 |
| OpenCode Zen | $0.3 | $2.5 | $0.03 |
| Pioneer | $0.3 | $2.5 | $0.03 |
| Requesty | $0.3 | $2.5 | $0.03 |
Rows tagged “first-party” are the model vendor's own pricing; untagged rows are resale channels and may differ.
When to choose it
- $0.3/$2.5 每百万 token 带 104 万上下文与图片输入 —— 要「便宜且能看图」时的少数选项
Caveats
- 输出单价($2.5)是输入的 8 倍多。任务如果是「短输入、长输出」(写长文、生成代码),实际账单会比按输入价估的高不少
Basis: first-party pricing page · reviewed 2026-08-09
Leaderboard
Not yet on any leaderboard we track. That does not mean the model performs poorly — it means we have not taken this snapshot for it yet, and we do not publish a score we have not measured.