Search the catalog, compare model details, and open the providers serving each model.

Hot and latest models
The first models are trending by recent visits, then the list fills with the newest releases in the catalog.
Latest news
A concise selection of recent announcements, releases, and independent coverage from across the catalog.
Fireworks AI
A July 21, 2026 benchmark from Fireworks AI compared Moonshot's open Kimi K3 model against Anthropic's Fable 5 on roughly 1,030 agent-based software development tasks, scoring Kimi K3 at 92.4% versus Fable 5 at 92.6%, placing the two within a fraction of a point of each other. On long-running agent processes the report For developers evaluating Fireworks AI's hosted Kimi K3 endpoint, the comparison implies near-parity agent coding performance at a substantially lower token cost for sustained workflows, though the underlying Fireworks benchmark page itself was not surfaced in the scraping set and is referenced only second-hand. The re
Vercel AI Gateway
MarkTechPost reports that the KwaiKAT Team has released KAT-Coder-V2.5, positioning it as an agentic coding model whose training regimen draws on more than 100,000 verifiable repository environments — a scale claim aimed at strengthening the model's ability to operate as an autonomous coding agent across realistic soft The piece is useful as third-party timing and framing for the underlying Kwaipilot release that powers the Vercel AI Gateway model, but the scraped excerpt is dominated by cookie consent banners and site advertising rather than substantive technical content. As a result, no concrete benchmarks, architecture details, or
Anthropic
CNBC reported on July 24, 2026 that Anthropic announced Claude Opus 5, describing it as the company's best-performing and most cost-effective model across industry benchmarks. The model outperforms the previously released Claude Fable 5 on coding and knowledge work evaluations, and Anthropic positioned it as "designed According to the CNBC report, Opus 5 is priced at $5 per million input tokens and $25 per million output tokens, exactly half the price of Fable 5, making it materially cheaper than the Mythos-tier model it edges out on coding and knowledge work. The launch lands in a more cost-conscious enterprise market where Anthrop
Anthropic released Claude Opus 5 on July 24, 2026, a new flagship model positioned for everyday business needs. The standout product feature is an adjustable effort toggle (low, medium, high) that lets users explicitly trade off cost against capability per request, addressing rising enterprise concerns about AI bills. According to the report, Opus 5 is Anthropic's fourth model release in less than two months, following Mythos 5, Fable 5, and Sonnet 5 in June 2026. The piece notes Fable 5 was the controversial model that the U.S. government temporarily placed export controls on after Amazon researchers reported they could bypass its
Inference
This IoT Digital Twin PLM explainer dated July 24, 2026 consolidates Gemma 3's lineage and deployment profile for developer audiences: it frames Gemma 3 (released March 2025) as a 27-billion-parameter open-weights model with native vision, a 128K-token context window, and quantization-aware-trained checkpoints that the The article is a third-party editorial retrospective, not an official Google or Inference.net source, and several of its claims (including a reported LMArena Elo around 1338 for the instruction-tuned 27B variant and the assertion that "Gemma 4 has shipped") are not corroborated by the supplied evidence set, so readers
Regolo AI
Regolo.ai published a case study on 2026-07-23 covering a University of Udine bachelor's thesis by Giovanni Battista Perini that selected Regolo as the language-model infrastructure for "Oracolo," a RAG-based conversational chatbot for citizens built on the Cheshire Cat framework and developed at the Treviso-based web The case study documents a deliberate architecture choice to address hallucination and stale-knowledge limitations of generative models, using Regolo as the model-access layer alongside retrieval-augmented generation. It serves as direct, provider-relevant evidence of third-party integration with api.regolo.ai, though
Providers
Open a provider page to see links, package names, and the models available through that provider.