Sulat.com
AI models
Requesty logo

Provider details

Requesty

Learn more about this provider, then browse the models currently listed under it.

requestyrequesty

Latest news about Requesty

Requesty

Official sourceAnnouncement

One-Stop Solution for AI Models — from the Requesty blog.

Requesty

Official sourceRelease Notes

Requesty's August 2026 changelog covers a dense set of platform upgrades for its unified LLM router. On August 19, the team improved Vertex AI auto caching so that cache creation is skipped when the cacheable prefix contains file attachments with non-gs:// URIs, while external URLs, inline/base64 files, and gs:// URIs Additional August 2026 releases include native video input over the Chat Completions API (supporting public video URLs and base64 data, with YouTube links translated to Gemini's native format), discounted-model visibility in the Model Library and a discount percentage field on the /models endpoint to surface DeepSeek o

Requesty

Coverage

PitchBook's company profile shows Requesty as a privately held, venture-backed London-based SaaS developer of a large-language-model gateway, founded in 2023 with roughly five employees. The company closed a $3.58M seed round on 26 September 2025, backed by four investors, and is described as generating revenue while s The profile frames Requesty within the broader AI infrastructure and enterprise SaaS venture landscape, pointing to reports such as the August 10, 2026 AI funding analysis showing $407B raised across megadeals and the August 3, 2026 Enterprise SaaS report on VC funding rebounding. Requesty's seed-round cap-table excerp

Requesty

Official sourceBenchmark

The Requesty model library lists GPT-5.5 on AWS Bedrock (us-east-1) as a managed endpoint, priced at $4.95 per 1M input and $29.70 per 1M output, reflecting a flat 10% discount versus the OpenAI list rates of $5.50 and $33.00. The endpoint exposes a 1M+ token context (922K input, 128K output) with vision, reasoning, to Workload cost examples on the page illustrate $0.79 for 100K input + 10K output, $7.92 for 1M input + 100K output, and $79.20 for 10M input + 1M output at those rates before caching savings, with cache reads reducing repeated-context costs further. The endpoint is invoked through Requesty's single OpenAI-compatible bas

About Requesty

What they do

Requesty provides a unified LLM gateway for routing, governance, and optimization across AI providers, aimed at enterprise control and spend management.

How they were founded

Founded by Daniel Trugman and Thibault Jaigu to help organizations govern and optimize AI model usage.

Quick Info

Organization
Requesty
Headquarters
London, England, United Kingdom
SDK package
@ai-sdk/openai-compatible
Synced at
Apr 24, 2026

Models served

142 models available through Requesty

These model results are sorted newest first so you can quickly see the latest options from this provider.

Requestyglm
GLM-5.3-Flash
glm-5.3-flashRequestyReleased Aug 26, 2026131,072 token context windowIn $0.075 · Out $0.25
Input
Output
Requestynemotron
nemotron-lightning-3.5-30b-a3b
nemotron-lightning-3.5-30b-a3bRequestyReleased Aug 15, 2026262,144 token context windowIn $0.05 · Out $0.20
Input
Output
Requestyglm
GLM-5.3 (EU)
glm-5.3@euRequestyReleased Aug 14, 20261,048,576 token context windowIn $1.20 · Out $4.20
Input
Output
Requestyglm
GLM-5.3
glm-5.3RequestyReleased Aug 14, 20261,048,576 token context windowIn $1.20 · Out $4.20
Input
Output
Requestygemini-flash
Gemini 3.7 Flash (EU)
gemini-3.7-flash@euRequestyReleased Aug 13, 20261,048,576 token context windowIn $0.825 · Out $4.125
Input
Output
Requestygemini-flash
Gemini 3.7 Flash
gemini-3.7-flashRequestyReleased Aug 13, 20261,048,576 token context windowIn $0.75 · Out $3.75
Input
Output
Requestyqwen
Qwen3.8 2.4T A95B
qwen3.8-2.4T-A95BRequestyReleased Aug 12, 2026262,144 token context windowIn $2.00 · Out $6.00
Input
Output
Requestygrok
Grok 4.6
grok-4.6RequestyReleased Aug 12, 2026500,000 token context windowIn $2.00 · Out $6.00
Input
Output
Requestydeepseek-thinking
DeepSeek V4 Pro 0813
deepseek-v4-pro-0813RequestyReleased Aug 12, 20261,000,000 token context windowIn $1.32 · Out $3.96
Input
Output
Requestynemotron
nemotron-3.5-lightning-30b-a3b
nemotron-3.5-lightning-30b-a3bRequestyReleased Aug 11, 20261,048,576 token context windowSubscription plan pricing
Input
Output
Requestymuse
Muse Glimmer 30B
muse-glimmer-30bRequestyReleased Aug 10, 2026131,072 token context windowSubscription plan pricing
Input
Output
Requestyling
ling-3.0-tiny
ling-3.0-tinyRequestyReleased Aug 5, 2026262,144 token context windowSubscription plan pricing
Input
Output
Requestyqwen
Qwen3.8 Max
qwen3.8-maxRequestyReleased Aug 3, 20261,048,576 token context windowIn $2.00 · Out $6.00
Input
Output
Requestydeepseek-flash
DeepSeek V4 Flash 0731 (EU)
deepseek-v4-flash-0731@euRequestyReleased Jul 31, 20261,048,576 token context windowIn $0.14 · Out $0.28
Input
Output
Requestydeepseek-flash
DeepSeek V4 Flash 0731
deepseek-v4-flash-0731RequestyReleased Jul 31, 20261,048,576 token context windowIn $0.076 · Out $0.153
Input
Output
Requestyclaude-opus
Claude Opus 5 (EU)
claude-opus-5@euRequestyReleased Jul 24, 20261,000,000 token context windowIn $5.50 · Out $27.50
Input
Output
Requestyclaude-opus
Claude Opus 5
claude-opus-5RequestyReleased Jul 24, 20261,000,000 token context windowIn $5.00 · Out $25.00
Input
Output
Requestygemini-flash-lite
Gemini 3.5 Flash Lite (EU)
gemini-3.5-flash-lite@euRequestyReleased Jul 21, 20261,048,576 token context windowIn $0.33 · Out $2.75
Input
Output
Previous pagePage 1 of 8