Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Venice AI logo

Model details

MiniMax M3 Preview

MiniMax M3 Preview is framed by NVIDIA's catalog listing as a multimodal Mixture-of-Experts vision-language model, pairing vision and language understanding with text, image, and video inputs while producing text outputs. The design emphasis on strong reasoning, coding, and tool-calling suggests a general-purpose assistant aimed at complex, multi-step workflows rather than narrow single-task use. The open-weights release, documented in third-party listings, positions it for teams that want to self-host or fine-tune rather than rely solely on a hosted endpoint. Together, these traits point to a model intended for research, prototyping, and applied work where both perception across modalities and deliberate reasoning matter.

In practice, the model's half-million-token context window and sizeable output budget make it well suited to long-document analysis, codebase reasoning, and agent-style tasks that interleave natural-language planning with tool calls. Its multimodal inputs allow pipelines that ingest diagrams, screenshots, or video frames alongside text, which is useful for technical documentation, UI understanding, and grounded question answering. The combination of explicit tool-calling support and reasoning focus makes it a natural fit for agentic applications, retrieval-augmented generation, and code-assistant scenarios where the model has to plan, invoke external functions, and synthesize results. Compared with purely text-only predecessors, the MoE vision-language architecture and tool-oriented capabilities represent a step toward more capable, general assistants that can both perceive and act.

Venice AIminimax-m3-previewminimax-m3

Quick Info

Powered by
Provider
Venice AI
Model key
minimax-m3-preview
Release date
Jun 12, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.20

Limits

Output tokens
65,536 tokens
Context window
524,288 tokens

Latest news about MiniMax M3 Preview

Venice AI

CoveragePreview

The TypingMind cost calculator provides concrete API specifications for the Venice-hosted MiniMax M3 Preview: a 524,288-token context window, 65,536 maximum output tokens, and pricing of $0.30 per million input tokens, $1.20 per million output tokens, and $0.06 per million cache-read tokens. It lists capabilities of at The page gives a release date of 2026-06-12 and a last-updated date of 2026-06-13, framed against Venice AI's privacy-focused positioning. Because TypingMind is a third-party aggregator, these specs should be treated as secondary corroboration rather than authoritative, but they are the only supplied candidate with har

Venice AI

Official sourcePreview

Venice AI lists MiniMax M3 Preview as an open-weights, natively multimodal large language model from MiniMax, hosted under a private zero-retention tier with TEE-based hardware enclaining and end-to-end encryption. The model uses a 428B-parameter Mixture-of-Experts architecture with 23B activated parameters and MiniMax For developers, the Venice-hosted M3 Preview is positioned for agentic coding, long-context research workflows, and cross-modal reasoning, with private inference guaranteeing that prompts and outputs are never stored or used for training. A third-party TypingMind pricing/spec mirror corroborates the 524,288-token conte

Videos about MiniMax M3 Preview

More models around MiniMax M3 Preview