Gemini 3.1 Flash-Lite
Google llmGoogle's most cost-effective Gemini 3 model, stated to be based on Gemini 3 Pro and optimised for high-volume, latency-sensitive tasks.
overall ⌄
36.3
67 of 93 ranked
price · $/M tokens
0.25 / 1.5
in / out, as of 2026-07-25
to run it yourself ⌄
API only
no weights to run
Benchmarks
What its maker published, and what anyone else measured. Every figure links the document it came from.
Cheaper, and at least as good
On a 1,000-in, 1,000-out call these cost less and rank no lower — a hard comparison to argue with, and a narrow one. Value covers what it misses.
| Devstral Small 2 Mistral AI | $0.00040 | overall 42.7 |
| Ministral 3 14B Reasoning Mistral AI | $0.00040 | overall 48.1 |
| DeepSeek V4-Flash DeepSeek | $0.00042 | overall 59.3 |
| DeepSeek V4-Pro DeepSeek | $0.00130 | overall 68.3 |
| MiniMax M2.1 MiniMax | $0.00150 | overall 50.9 |
Lineage
↑ distill Gemini 3 Pro
↓ finetune Gemini 3.5 Flash-Lite
Metadata
| org | |
| released | 2026-03-03 |
| license | proprietary |