Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Requesty logo

Model details

Laguna XS.2

Laguna XS.2 is a text-to-text model from poolside built around a Mixture-of-Experts design with roughly 33B total parameters but only about 3B activated per token, a sparsity profile that keeps compute and memory demands in check while still supporting larger-capacity reasoning paths. The architecture mixes Sliding Window Attention with a global attention layout, applying sigmoid-gated per-head routing in 30 of the 40 transformer layers to limit KV cache growth and speed up inference on a single workstation. Poolside developed it for agentic coding and other long-horizon work that benefits from running locally, and the weights are openly published so developers can self-host the model with quantization or run it through a hosted endpoint.

The model targets practical developer workflows rather than broad multimodal coverage: it accepts and returns text only, supports tool calling for agent loops, and ships with guardrail controls for safer integration into pipelines. A third-party feature listing reports a context input size around 131.1K tokens, which gives it room to hold sizeable code repositories, multi-file refactors, or extended task histories in a single prompt. Its combination of small active-parameter footprint, open weights, and agent-oriented capabilities makes it a reasonable fit for teams who want to run coding assistants on their own hardware while still being able to orchestrate tools and longer-running tasks without offloading everything to a remote service.

Requestylaguna-xs.2laguna

Quick Info

Powered by
Provider
Requesty
Model key
laguna-xs.2
Release date
Apr 28, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
32,768 tokens
Context window
32,768 tokens

Latest news about Laguna XS.2

Requesty

Coverage

A Techmeme aggregator snapshot dated April 29, 2026 surfaces VentureBeat reporting (by Carl Franzen) that US startup Poolside debuted its first open-weight model, Laguna XS.2, described as a 33B-total/3B-activated Mixture-of-Experts architecture. The same item notes Poolside simultaneously released Laguna M.1, a separa The headline-level news here is a formal model debut rather than a serving update: Poolside's first open-weight release, with the XS.2 positioned as the publicly accessible member of a new Laguna family that also includes the larger proprietary M.1 sibling. No API pricing, routing changes, or gateway availability claim

Requesty

Official sourceBenchmark

Requesty lists Poolside's Laguna XS.2 as a managed chat endpoint under the model id "poolside/laguna-xs.2", accessible via the OpenAI-compatible base URL https://router.requesty.ai/v1. The deployment was added in June 2026, is served from the US, and exposes a 33K token context window with chat as its sole capability ( The upstream Poolside rates shown on the Requesty page are free for both input and output per 1M tokens, with no per-request fee, though Requesty's pay-as-you-go markup (0% when using your own keys, 5% otherwise) applies on top. The page does not publish benchmark scores for this specific variant, directing developers

Videos about Laguna XS.2

More models around Laguna XS.2