Currently listed through these providers:
Model details
Ornith 1.5 9B Thinking
Ornith 1.5 9B Thinking sits inside the broader Ornith 1.5 family and is the smaller, efficiency-oriented sibling of the lineup, paired with an explicit "thinking" mode that favors more deliberate reasoning rather than fast first-token replies. The NanoGPT listing frames the variant as purpose-built for agentic coding, tool use, visual understanding, and long-context workloads, signaling that the lab tuned it for workflows where the model has to plan, call tools, parse images, and stay coherent across very large documents. That positioning makes it a practical middle ground for teams that want a capable reasoning model without paying the much higher cost of the 397B Thinking variant in the same family, which the aggregator source prices at roughly $0.9 per million input and $3.60 per million output tokens.
In day-to-day use, the model is well suited to coding assistants that need to invoke external tools and read attached images, document-heavy tasks where a 262K-token window matters, and structured-output pipelines that benefit from the thinking phase producing more carefully checked answers. Because it is open weights and ships in FP8+ serving configurations on NanoGPT, it is also a reasonable choice for self-hosting or cost-sensitive production deployments where predictable per-million-token pricing and cache reuse matter. The official page does not publish benchmark numbers, so any concrete performance claim should come from a follow-up evaluation rather than the vendor copy, but the combination of multimodal input, tool calling, structured output, and a very large context window gives it a flexible profile for builders assembling agent-style applications.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- ornith-ai/ornith-1.5-9b:thinking
- Release date
- Jul 29, 2026
- Last updated
- Aug 24, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.20
Limits
- Input tokens
- 262,144 tokens
- Output tokens
- 32,768 tokens
- Context window
- 262,144 tokens