MiniMax M2

llm

A 230B-parameter (10B active) MoE model that deliberately returned to full attention over M1's linear attention, optimized for coding and agentic workflows.

Metadata

orgMiniMax
released2025-10-27
params230B-A10B
licensemodified-mit
kindllm
hf idMiniMaxAI/MiniMax-M2

Benchmarks

MMLU-Pro82source ↗
GPQA78source ↗
AIME 202578source ↗
SWE-bench69.4source ↗

Lineage

↓ finetuneMiniMax M2.1
view in full graph →