Currently listed through these providers:
Model details
o4 Mini High
o4 Mini High sits inside OpenAI's "o" reasoning family as a compact but capability-leaning member, positioned alongside the larger o3 and the standard o4-mini variants. The design intent of the family is to combine careful chain-of-thought reasoning with multimodal understanding and full access to tools such as web browsing, code execution, and structured outputs, so the model can think step by step about images, text, and documents rather than just produce a quick answer. Within that lineup, the "high" tier signals a configuration that trades extra compute for deeper deliberation per response, making it suited to tasks where careful reasoning over visual and textual inputs matters more than raw throughput, such as analyzing complex diagrams, multi-step scientific problems, or workflows that chain together reasoning and tool calls in a single trajectory.
The lineage of o4 Mini High traces directly to OpenAI's o-series reasoning research, following the path laid out by o1 and then the o3 generation that early commentary described as the company's most capable models yet at the time of release. Sources characterize the o3 and o4-mini generation as bringing "thinking with images" and richer tool use into the same model, and o4 Mini High inherits that multimodal-reasoning hybrid approach while remaining a smaller, faster sibling to o3. In practical terms this means it is shaped to act as a reasoning workhorse that can take in images, PDFs, and text, plan across many steps with tool calling and structured output, and return carefully worked-out answers within a very large context window, fitting naturally into agent-style pipelines where a lighter model still needs to plan, verify, and reflect before responding.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- openai/o4-mini-high
- Release date
- Apr 16, 2025
- Last updated
- Apr 16, 2025
- Knowledge cutoff
- 2024-06-30
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.10
- Output token cost
- $4.40
Limits
- Output tokens
- 100,000 tokens
- Context window
- 200,000 tokens