Claude 3.7 Sonnet
llmAnthropic's first hybrid reasoning model, letting users toggle standard and extended-thinking modes. Anthropic publishes no single-attempt extended-thinking GPQA figure for this model, so the standard-mode score is recorded.
Metadata
| org | Anthropic |
| released | 2025-02-24 |
| license | proprietary |
| kind | llm |
Benchmarks
| GPQA | 68 | source ↗ |
| SWE-bench | 63.7 | source ↗ |
Lineage
No recorded lineage — root or standalone model.
view in full graph →