Sulat.com
AI models
NanoGPT logo

Model details

MiniMax M3

MiniMax M3 is framed by its distribution partners as a coding and agentic foundation model with native multimodality, accepting both text and image inputs while producing text responses. The Ollama library listing brands it as a "Coding & Agentic Frontier" release, advertising a one-million-token context framing alongside a native multimodal design that is marketed for tool-using workflows. A community thread on the NVIDIA developer forums shows enthusiasts already exploring NVFP4 quantization paths for the model on a Quad DGX Spark configuration, which signals active interest in running M3 in latency-sensitive, on-prem inference setups rather than only through managed endpoints.

In practice, the model is presented as a strong fit for autonomous coding assistants, agentic pipelines, and long-context retrieval tasks where interleaved text and image reasoning matter. The Ollama distribution lists vision, tools, and thinking capability toggles, suggesting it is wired for structured tool calling and stepped reasoning rather than pure chat, and the official cloud listing emphasizes commercial licensing with zero data retention for teams that need a managed deployment. The community quantization work, combined with the multimodal and agent-oriented marketing, points to a model intended for developers who want frontier-style reasoning plus the option to self-host quantized variants on compact accelerators when data sovereignty or cost control is a priority.

NanoGPTminimax/minimax-m3minimax

Quick Info

Powered by
Provider
NanoGPT
Model key
minimax/minimax-m3
Release date
Jun 1, 2026
Last updated
Jun 1, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.20

Limits

Input tokens
512,000 tokens
Output tokens
80,000 tokens
Context window
512,000 tokens

Transparent token rates

Compare MiniMax M3 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiniMax M3

Vercel AI Gateway

Official sourceAnnouncement

MiniMax Research officially released MiniMax M3 on June 1, 2026, positioning it as the first and only open-weight model to combine frontier-level coding/agentic performance, native multimodality (image and video input), desktop-computer operation, and ultra-long context in a single model. The post claims significant co The release introduces MSA (MiniMax Sparse Attention), a new sparse attention architecture proposed by MiniMax's team that underpins the model's 1M-token context window by addressing the quadratic complexity of full attention. The blog details M3's availability through MiniMax Code, the Token Plan, and the MiniMax API,

Videos about MiniMax M3

More models around MiniMax M3