Claude 3.7 Sonnet

llm

Anthropic's first hybrid reasoning model, letting users toggle standard and extended-thinking modes. Anthropic publishes no single-attempt extended-thinking GPQA figure for this model, so the standard-mode score is recorded.

Metadata

orgAnthropic
released2025-02-24
licenseproprietary
kindllm

Benchmarks

GPQA68source ↗
SWE-bench63.7source ↗

Lineage

No recorded lineage — root or standalone model.

view in full graph →