Friendli
FriendliAI opens a 7,000 sq ft San Francisco office in SoMa to scale frontier AI inference for open-weight and custom models as agentic workloads surge.
Provider details
Learn more about this provider, then browse the models currently listed under it.
Friendli
FriendliAI opens a 7,000 sq ft San Francisco office in SoMa to scale frontier AI inference for open-weight and custom models as agentic workloads surge.
Friendli
March 13, 2026 — FriendliAI has launched Friendli InferenceSense, an inference monetization platform purpose-built for GPU cloud operators.
Friendli
Friendli InferenceSense - the "AdSense for GPUs" - Monetize Idle GPU Capacity with Inference "Assembly Line" for AI Factories
Friendli
FriendliAI's official changelog documents an active Model APIs lifecycle through July 2026, with developers seeing frequent model additions and deprecations on the serverless platform. Recent dated entries include the deprecation of the Tool Assisted API on July 6, 2026, the removal of zai-org/GLM-5 on July 3, 2026, an Earlier in 2026 the changelog also surfaces removals such as deepseek-ai/DeepSeek-V3.1 (May 22), Qwen/Qwen3-30B-A3B (April 16), zai-org/GLM-4.7 (April 15), and MiniMaxAI/MiniMax-M2.1 (April 9), showing a steady rotation of open-weight model coverage on FriendliAI's Model APIs. The changelog UI additionally exposes filt
What they do
Friendli provides serverless inference endpoints and deployment tooling for generative AI models, with APIs for hosted open and partner models.
How they were founded
Founded in 2021 by Byung-Gon Chun to make deployment and serving of generative AI models easier for developers and enterprises.
Models served
These model results are sorted newest first so you can quickly see the latest options from this provider.