Add corrections, implementation notes, pricing changes, or usage caveats for other readers.
Last updated
Jun 12, 2026
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities
262,000 tokens
Recent tweets and retweets from Baseten
Open-source models have reached a new level of intelligence; our own engineers use models like GLM-5.2 for tasks like kernel optimization.
Thanks @latentspacepod for having @waterloo_intern and @philipkiely on to talk about how we optimize models for the highest throughput…
One of our engineers asked @poolsideai's Laguna S 2.1 to transform a 715-file C++ game from neon cyberpunk into an Ancient Greek aesthetic.
It orchestrated 3 different models, refactored the code, and produced a playable build.
baseten.co/blog/laguna-s-21-…
Try DeepSeek V4 Flash on our Model APIs today.
- 80–98% cheaper than other frontier models
- Comparable intelligence
- 1M token context window
Use it here: baseten.co/library/deepseek-…
Discuss this model
Add corrections, implementation notes, pricing changes, or usage caveats for other readers.