Alibaba (China)
MiniMax announced M2.7 as the first model in its line that deeply participated in its own evolution, using agent teams, complex skills, and dynamic tool search to update its own memory and build dozens of harness skills for reinforcement learning experiments. The release emphasizes self-evolution, where M2.7 improves its learning process and harness based on experiment results. This positions the model for highly elaborate productivity tasks beyond earlier M2-series iterations. Benchmark results reported by MiniMax include 56.22% on SWE-Pro, 55.6% on VIBE-Pro, and 57.0% on Terminal Bench 2, alongside a GDPval-AA ELO of 1495. The model shows improved complex editing in Office suite applications including Excel, PowerPoint, and Word, with better handling of multi-turn modifications and high-fidelity edits. M2.7 also demonstrates stronger identity preservation and emotional intelligence, expanding its use into broader interactive entertainment scenarios.