Gemini 3.6 Flash

google · google/gemini-3.6-flash

Fast Gemini model balancing multimodal reasoning, tool use, and cost

≈699k Chinese characters of context · Reads images · Step-by-step reasoning · Calls tools / agents · Structured JSON output

Read a 200,000-Chinese-character book and write a 2,000-Chinese-character summary: about $0.24

An order-of-magnitude estimate at 1.5 tokens per Chinese character and the vendor's own rate — not a quote. Full method: the glossary on the AI Models page.

Specs

Context1,048,576(≈699k Chinese characters)
Max output65,536
Structured outputYes
Open weightsNo
Released2026-07-21

Source: models.dev snapshot 2026-08-29

Who serves it, at what price

ProviderIn /MOut /MCache read /M
Googlefirst-party$0.75$3.75$0.075
Abacus$1.5$7.5$0.15
Cortecs$0.75$3.75$0.075
GitHub Copilot$0.75$3.75$0.075
Vertex$0.75$3.75$0.075
DevPass (LLM Gateway)$0.75$3.75$0.075
OpenCode Zen$1.5$7.5$0.15
Pioneer$1.5$7.5$0.15
Requesty$1.5$7$0.15

Rows tagged “first-party” are the model vendor's own pricing; untagged rows are resale channels and may differ.

When to choose it

This model is on our watch list but has no written assessment yet — we do not write what we cannot support. The specs and prices above still stand.

Leaderboard

Not yet on any leaderboard we track. That does not mean the model performs poorly — it means we have not taken this snapshot for it yet, and we do not publish a score we have not measured.