GMI Cloud
Artificial Analysis published a detailed independent evaluation of MiniMax-M3 (article dated June 8, 2026), positioning it as MiniMax's first multimodal M-series model with a 1M-token context window and native vision input (image and video), succeeding the text-only MiniMax-M2.7. The model scores 55 on the Artificial A The article documents specific capability gains over MiniMax-M2.7: HLE +9 points (28% to 37%), GPQA Diamond +6 (87% to 93%), AA-LCR +5 (69% to 74%), IFBench +7 (76% to 83%), and CritPt +3 (1% to 4%), with a small SciCode regression (47% to 45%). It scores 1670 on GDPval-AA (level with Claude Sonnet 4.6 max), 80% on MMM