Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Inco logo

Provider details

Inco

Learn more about this provider, then browse the models currently listed under it.

incoinco
Provider docs

Latest news about Inco

Inco

Official sourceAnnouncement

The Inco Decode blog index lists three recent first-party announcements: the September 3, 2026 launch of the Inco platform on Artificial Analysis (covering Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash); an August 28, 2026 post on day-0 support for GLM 5.3 alongside DFlash 2 and NVFP4 checkpoints for faster, more eff The index confirms the chronology of Inco's stack evolution across late summer 2026, framing DFlash 2 as the underlying speculative-decoding technique and the GLM 5.3 day-0 launch as an application of that technique combined with NVFP4 quantization. The reference to "Muse Glimmer" appears only in this index entry and i

Inco

Official sourceAnalysis

On September 3, 2026, Inco AI announced the public-beta launch of the Inco inference platform, releasing high-speed endpoints for Kimi K3, MiniMax M3, GLM 5.3, and GLM 5.3 Flash. According to the post, each model leads its respective Artificial Analysis provider leaderboard on output speed as captured on September 8, 2 The post details four stack layers that Inco credits for the performance results: (1) fleet-scale serving with cache-aware routing and scheduling; (2) speculative decoding via DFlash and DFlash 2, parallel block-diffusion techniques the post attributes to the Inco team, with open-source checkpoints already available; (

Inco

Coverage

The incoai Hugging Face organization page lists Inco AI's published open-source artifacts, including DFlash and DFlash 2 text-generation checkpoints (1B, 2B, and 3B sizes), as well as DFlash 2 drafters for Qwen3.8-27B and GLM-5.3-Flash, plus GGUF variants. The page also shows a 391B text-generation model entry. These p Recent activity on the organization page shows updates from a team member (zhijianliu) within the last several hours and days, including updates to incoai/Qwen3.8-27B-DFlash2 (about 3 hours ago at capture) and incoai/GLM-5.3-Flash-DFlash2 (18 days ago), alongside a GLM 5.3 Flash DFlash 2 drafter updated 20 days ago. Th

About Inco

What they do

Inco AI operates a prepaid, pay-per-token inference API compatible with the OpenAI and Anthropic SDKs, serving optimized endpoints for open-weight models such as DeepSeek V4.1 Flash, Kimi K3, MiniMax M3, and GLM 5.3.

How they were founded

Public sources reviewed do not name individual founders. Inco AI's LinkedIn page lists a 2026 founding and Palo Alto headquarters, and the company's blog says its platform entered public beta in September 2026, building on DFlash speculative-decoding research seeded at Z Lab.

Quick Info

Organization
Inco AI
Headquarters
Palo Alto, California, United States
SDK package
@ai-sdk/openai-compatible
Synced at
Sep 18, 2026

Models served

7 models available through Inco

These model results are sorted newest first so you can quickly see the latest options from this provider.

Incodeepseek-flash
DeepSeek V4.1 Flash
deepseek-v4.1-flash:fastIncoReleased Sep 10, 20261,000,000 token context windowIn $0.60 · Out $2.40
Input
Output
Incoglm-flash
GLM-5.3-Flash
glm-5.3-flash:fastIncoReleased Aug 26, 20261,000,000 token context windowIn $0.15 · Out $0.50
Input
Output
Incoglm
GLM-5.3 Fast
glm-5.3:fastIncoReleased Aug 14, 20261,000,000 token context windowIn $2.80 · Out $8.80
Input
Output
Incoglm
GLM-5.3
glm-5.3IncoReleased Aug 14, 20261,000,000 token context windowIn $1.40 · Out $4.40
Input
Output
Incokimi-k3
Kimi K3
kimi-k3:fastIncoReleased Jul 16, 20261,048,576 token context windowIn $6.00 · Out $30.00
Input
Output
Incominimax
MiniMax M3 Fast
minimax-m3:fastIncoReleased Jun 1, 20261,048,576 token context windowIn $0.60 · Out $2.40
Input
Output
Incominimax
MiniMax-M3
minimax-m3IncoReleased Jun 1, 20261,048,576 token context windowIn $0.30 · Out $1.20
Input
Output