Friendli
The Anthropic Messages API is now supported on FriendliAI. Run open models like DeepSeek to cut costs without rewriting code, plus get up to $50k in credits.
Model details
MiniMax M2.5 was introduced on February 12, 2026 as part of the MiniMax LLM family, sitting alongside the M2.7 and M3 models on the vendor's text-models page where it is positioned as state-of-the-art for coding and agent tasks and "Designed for Agent Universe." The model is built on a Mixture-of-Experts architecture with 230 billion total parameters but only about 10 billion active during any given inference pass, a sparse-activation design that the vendor and third-party reviewers credit with keeping serving costs low while still targeting frontier-class reasoning quality. By keeping most of the network idle per token, the MoE layout lets M2.5 aim at complex agent workloads without paying full dense-model compute on every request.
M2.5 is aimed squarely at developers building agentic coding assistants, tool-using workflows, and long-context automation. The vendor page frames it as a coding- and agent-focused release, and it ships within an ecosystem that lets teams run it through the MiniMax API or self-host it, giving engineering teams flexibility in how they deploy. FriendliAI's inference platform has added Anthropic Messages API compatibility for open models, which is the route through which this model can be served with minimal integration changes for teams already structured around that protocol. The combination of sparse activation, a large effective context, and agent-oriented positioning makes M2.5 a practical fit for code-generation pipelines, multi-step tool orchestration, and other developer workflows where both reasoning depth and economical per-token serving matter.
Friendli
The Anthropic Messages API is now supported on FriendliAI. Run open models like DeepSeek to cut costs without rewriting code, plus get up to $50k in credits.
Friendli
MiniMax's official platform release-notes index lists the Feb 2026 entry: "Released MiniMax-M2.5 series models (MiniMax-M2.5 / M2.5-highspeed). Achieving or setting new SOTA benchmarks in programming, tool calling and search, office productivity and other scenarios." This is first-party confirmation of the model line, The same index contextualizes M2.5 within MiniMax's broader 2026 model roadmap, including later releases such as MiniMax-M2.7 (Mar 18, 2026), MiniMax M3 (Jun 1, 2026), and MiniMax H3 (Jul 31, 2026), alongside music and speech updates. Together these entries confirm M2.5 as a stepping stone in MiniMax's iterative M-seri
Friendli
OpenRouter's official provider page lists MiniMax-M2.5 with a 205K context window, $0.30/M input and $1.20/M output listed pricing, and a released date of Feb 12, 2026. The model is described as a state-of-the-art LLM trained on complex real-world digital working environments, building on M2.1's coding strengths to add Friendli is shown as one of nine available providers for MiniMax-M2.5 on OpenRouter, with listed pricing of $0.30 input / $1.20 output per 1M tokens, a $0.06 cache-read rate, 0.13s P50 latency, and 188 tokens per second throughput — the highest throughput and lowest latency among the listed providers. Other providers i