Gemini 3.1 Flash-Lite

Google llm

Google's most cost-effective Gemini 3 model, stated to be based on Gemini 3 Pro and optimised for high-volume, latency-sensitive tasks.

overall ⌄

36.3

67 of 93 ranked

price · $/M tokens

0.25 / 1.5

in / out, as of 2026-07-25

to run it yourself ⌄

API only

no weights to run

Benchmarks

What its maker published, and what anyone else measured. Every figure links the document it came from.

benchmarkpublishedmeasured
GPQA86.9
SWE-bench Pro38.3
HLE16

Cheaper, and at least as good

On a 1,000-in, 1,000-out call these cost less and rank no lower — a hard comparison to argue with, and a narrow one. Value covers what it misses.

Devstral Small 2 Mistral AI$0.00040overall 42.7
Ministral 3 14B Reasoning Mistral AI$0.00040overall 48.1
DeepSeek V4-Flash DeepSeek$0.00042overall 59.3
DeepSeek V4-Pro DeepSeek$0.00130overall 68.3
MiniMax M2.1 MiniMax$0.00150overall 50.9

Lineage

↑ distill Gemini 3 Pro
view in full graph →

Metadata

orgGoogle
released2026-03-03
licenseproprietary