Currently listed through these providers:
Model details
Claude Haiku 4.5
Claude Haiku 4.5 is the lightweight, speed-optimized member of the Haiku line, designed to deliver near-frontier intelligence in a form factor that prioritizes low latency and high throughput over raw capacity. Where larger Claude models chase the absolute ceiling of reasoning, Haiku 4.5 is built for real-time, high-volume settings where responsiveness and cost dominate, making it well suited to chat assistants, sub-agents handling parallel work, and scaled production deployments. The design intent is to keep the model compact and fast without giving up the reasoning, coding, and tool-use quality that defines the broader Claude family, and to do so at a price point that makes always-on use economically sensible. A notable design addition is the introduction of extended thinking to the Haiku line, giving the model controllable reasoning depth with summarized or interleaved thought output, and full support for tool-assisted workflows including coding, bash, web search, and computer-use tools. This lets a smaller, cheaper model behave less like a thin responder and more like a capable agent that can plan, call tools, and act on results. In practice, Haiku 4.5 is positioned as the workhorse of the Claude family: the model you route to when you need fast, reliable, agentic behavior at scale, while reserving larger models for the hardest problems that justify their extra cost and latency.
On qualitative strengths, Haiku 4.5 matches the previous-generation Claude Sonnet 4 across reasoning, coding, and computer-use tasks, which is the central pitch: a model one tier down in size and cost can stand in for flagship-class behavior on the workloads most teams actually run. That claim is backed by a 73% score on SWE-bench Verified, placing it among the strongest coding models available and reinforcing its fit for agentic coding, refactoring assistance, and tool-driven development. The combination of frontier-adjacent coding ability, extended thinking, and broad tool support makes it attractive as a sub-agent that handles well-scoped tasks in parallel, freeing a larger orchestrating model to focus on planning and synthesis. For forward-looking use, Haiku 4.5 is shaped for workflows that lean on many small, fast model calls rather than a single heavyweight inference: large-scale code review, retrieval-augmented assistants, customer-facing chat, and agent fleets where dozens of instances must respond in near real time. The economics reinforce that posture, with standard input and output rates that are a fraction of larger Claude models, encouraging developers to use it as a default high-throughput option and to escalate only when a task demands more. The takeaway is a model whose lineage points toward cheap, fast, agentic intelligence that can be deployed at scale, and whose capability profile is broad enough to take on a meaningful share of the work usually reserved for flagship systems.
Quick Info
Powered by- Provider
- OpenCode Zen
- Model key
- claude-haiku-4-5
- Release date
- Oct 15, 2025
- Last updated
- Oct 15, 2025
- Knowledge cutoff
- 2025-02-28
- AI SDK package
@ai-sdk/anthropic- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.00
- Output token cost
- $5.00
Limits
- Output tokens
- 64,000 tokens
- Context window
- 200,000 tokens
Latest news about Claude Haiku 4.5
Videos about Claude Haiku 4.5
More models around Claude Haiku 4.5
This exact model name is also listed by 40 other providers.