Currently listed through these providers:
Model details
claude-haiku-4-5-20251001
Claude Haiku 4.5 is the compact, speed-oriented member of the Claude 4.5 generation, designed to deliver capable language and vision understanding at a lower cost than its larger siblings. It accepts text, image, and PDF inputs and produces text and tool-call outputs, making it well suited to multimodal pipelines where documents, screenshots, or diagrams need to be parsed alongside natural language instructions. The model is built for production assistants, coding agents, support tools, and structured reasoning workflows, with a wide context window and a generous output ceiling that allow it to hold long documents, multi-turn conversations, and extended agent traces in memory without constant truncation. Within the Claude family, Haiku 4.5 is positioned as the high-throughput workhorse: it is intended to feel responsive in interactive settings while still benefiting from the same multimodal and tool-use improvements introduced across the 4.5 series. Its combination of vision, structured output, streaming, and reliable function calling makes it a natural fit for retrieval-augmented chat, customer support automation, lightweight code generation, and any scenario where teams want the Claude quality bar in a form that can be called frequently and at scale. The balance of capability and efficiency suggests a model aimed squarely at everyday production traffic rather than heavyweight frontier tasks.
As part of Anthropic's staggered Claude 4.5 rollout, Haiku 4.5 inherits the architectural lineage of the series while being optimized for the low-latency, high-volume tier of the lineup, and it is offered both through Anthropic's own API and through Amazon Bedrock, giving deployers a choice of hosting environment with matching pricing and capability profiles. Its feature set mirrors the broader family, supporting tool calling, tool choice, structured output, and streaming alongside vision input, so it can plug into existing Claude 4.5 agent frameworks without special handling. The model is available on managed routing providers that pass through the official price points, which keeps the cost economics of Haiku 4.5 predictable for builders composing multi-model systems. For forward-looking use, Haiku 4.5 is best understood as the everyday runtime layer of the Claude 4.5 stack: strong enough to handle coding copilots, autonomous agent loops, document Q&A, and customer-facing chat, while remaining affordable enough to sit behind always-on features in consumer products. Teams can pair it with larger Claude 4.5 models for complex planning and then hand off execution, summarization, and tool orchestration to Haiku 4.5, getting a clean division of labor between depth and responsiveness within a single model family.
Quick Info
Powered by- Provider
- 302.AI
- Model key
- claude-haiku-4-5-20251001
- Release date
- Oct 16, 2025
- Last updated
- Oct 16, 2025
- Knowledge cutoff
- 2025-02-28
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.00
- Output token cost
- $5.00
Limits
- Output tokens
- 64,000 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare claude-haiku-4-5-20251001 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about claude-haiku-4-5-20251001
No articles yet. Fetch the latest news to show it here.