SIDE BY SIDE
Compare selected AI models side by side.
Compare uses, prices, speed evidence, and version details for up to three models. Language, video, decision models and historical evaluations are shown separately.
Language & coding
| What matters | Gemini 3.8 Flash (fast) ↗Google |
|---|---|
| Use cases | Coding · Agents & tools · Reasoning |
| Model score | 41Gemini 3.8 Flash · highArtificial Analysis · Gemini 3.8 Flash · high ↗ |
| Output / 1M tokens · USD | $0.75 |
| Published rates · USD | $0.75 input / $3.75 outputPer 1M native tokens |
| Cached input / 1M | $0.075 |
| Cache write / 1M | Not verified |
| Pricing conditions | Gemini Developer API · Standard · introductoryThese rates end December 31, 2026. Input/output then become $1.50/$7.50 per MTok. Cache storage is extra; output includes thinking tokens.Gemini Developer API pricing ↗ |
| Output speed · tokens/s | 236.5 tokens/sGoogle API · Gemini 3.8 Flash · High reasoning. Artificial Analysis reports a rolling 72-hour median of output speed after the first token, not total response time. Checked 2026-10-05; speed varies with workload and serving conditions.Artificial Analysis · Gemini 3.8 Flash (High) ↗Artificial Analysis · performance methodology v2.2.0 ↗ · Checked 2026-10-05 |
| Context window | 1,000,000 |
| Model ID | gemini-3.8-flash |
| Version notes | The Flash name alone is not used as evidence of measured speed. |
| Official documentation | Gemini 3.8 Flash ↗Checked 2026-10-05 |
Language prices are USD per million tokens; input and output are billed separately. Video samples use 5s of 768P output where verified. Extra tools, reasoning, cache writes, input media, storage, and taxes can change the total. (fast) marks a published median output speed above 150 tokens/s in the linked test settings.