← Choose models

Language & coding

What mattersGemini 3.8 Flash (fast) ↗Google
Use casesCoding · Agents & tools · Reasoning
Model score41Gemini 3.8 Flash · highArtificial Analysis · Gemini 3.8 Flash · high ↗
Output / 1M tokens · USD$0.75
Published rates · USD$0.75 input / $3.75 outputPer 1M native tokens
Cached input / 1M$0.075
Cache write / 1MNot verified
Pricing conditionsGemini Developer API · Standard · introductoryThese rates end December 31, 2026. Input/output then become $1.50/$7.50 per MTok. Cache storage is extra; output includes thinking tokens.Gemini Developer API pricing ↗
Output speed · tokens/s236.5 tokens/sGoogle API · Gemini 3.8 Flash · High reasoning. Artificial Analysis reports a rolling 72-hour median of output speed after the first token, not total response time. Checked 2026-10-05; speed varies with workload and serving conditions.Artificial Analysis · Gemini 3.8 Flash (High) ↗Artificial Analysis · performance methodology v2.2.0 ↗ · Checked 2026-10-05
Context window1,000,000
Model IDgemini-3.8-flash
Version notesThe Flash name alone is not used as evidence of measured speed.
Official documentationGemini 3.8 Flash ↗Checked 2026-10-05
Language prices are USD per million tokens; input and output are billed separately. Video samples use 5s of 768P output where verified. Extra tools, reasoning, cache writes, input media, storage, and taxes can change the total. (fast) marks a published median output speed above 150 tokens/s in the linked test settings.