Gemini 3.5 Flash Lite

google · google/gemini-3.5-flash-lite

Fast Gemini model balancing multimodal reasoning, tool use, and cost

≈699k Chinese characters of context · Reads images · Step-by-step reasoning · Calls tools / agents · Structured JSON output

Read a 200,000-Chinese-character book and write a 2,000-Chinese-character summary: about $0.10

An order-of-magnitude estimate at 1.5 tokens per Chinese character and the vendor's own rate — not a quote. Full method: the glossary on the AI Models page.

Specs

Context1,048,576(≈699k Chinese characters)
Max output65,536
Structured outputYes
Open weightsNo
Released2026-07-21

Source: models.dev snapshot 2026-08-29

Who serves it, at what price

ProviderIn /MOut /MCache read /M
Googlefirst-party$0.3$2.5$0.03
Abacus$0.3$2.5$0.03
Cortecs$0.33$2.749$0.033
Vertex$0.3$2.5$0.03
DevPass (LLM Gateway)$0.3$2.5$0.03
OpenCode Zen$0.3$2.5$0.03
Pioneer$0.3$2.5$0.03
Requesty$0.3$2.5$0.03

Rows tagged “first-party” are the model vendor's own pricing; untagged rows are resale channels and may differ.

When to choose it

  • $0.3/$2.5 每百万 token 带 104 万上下文与图片输入 —— 要「便宜且能看图」时的少数选项

Caveats

  • 输出单价($2.5)是输入的 8 倍多。任务如果是「短输入、长输出」(写长文、生成代码),实际账单会比按输入价估的高不少

Basis: first-party pricing page · reviewed 2026-08-09

Leaderboard

Not yet on any leaderboard we track. That does not mean the model performs poorly — it means we have not taken this snapshot for it yet, and we do not publish a score we have not measured.