Claude 3.7 Sonnet
Anthropic llmAnthropic's first hybrid reasoning model, letting users toggle standard and extended-thinking modes. Anthropic publishes no single-attempt extended-thinking GPQA figure for this model, so the standard-mode score is recorded.
overall ⌄
35.9
70 of 93 ranked
price · $/M tokens
no published price
to run it yourself ⌄
API only
no weights to run
Benchmarks
What its maker published, and what anyone else measured. Every figure links the document it came from.
Lineage
No recorded lineage — root or standalone model.
Metadata
| org | Anthropic |
| released | 2025-02-24 |
| license | proprietary |