Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Laguna XS 2.1

Laguna XS 2.1 is poolside's open-weight coding agent model built around a Mixture-of-Experts design, pairing 33 billion total parameters with only about 3 billion active per token so the heavy lifting is gated to specialized experts while keeping inference lean. That ratio is what allows the model to be quantized down to roughly 16–20 GB of VRAM at INT4, putting real local agentic coding within reach of a workstation rather than a data center. The release is positioned as a direct successor to the earlier Laguna XS.2, refining the family rather than reinventing it, and it continues poolside's broader push into open models paired with a runtime that supports training and serving agents.

On practical workloads, the upgrade over XS.2 shows up most clearly in software engineering evaluations and terminal-style tasks, where poolside reports a 5.4 point gain on SWE-bench Multilingual alongside stronger terminal benchmarks and an expanded 262K context window that comfortably handles long agent traces and multi-file refactors. The model is aimed squarely at agentic coding and long-horizon work on a local machine, which makes it a natural fit for developers who want a self-hosted reasoning and tool-using partner without depending on a hosted frontier API. Its combination of open weights, MoE efficiency, and benchmark gains over the previous Laguna release makes it one of the more compelling options in the open-weight coding agent space.

OpenRouterpoolside/laguna-xs-2.1laguna

Quick Info

Powered by
Provider
OpenRouter
Model key
poolside/laguna-xs-2.1
Release date
Jul 2, 2026
Last updated
Jul 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.06
Output token cost
$0.12

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Transparent token rates

Compare Laguna XS 2.1 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Laguna XS 2.1

LLMTR

Official sourceAnnouncement

Poolside officially introduced Laguna XS 2.1 on July 2, 2026, as an upgraded version of its Laguna XS.2 coding model. Per the company's blog post, Laguna XS 2.1 is a 33B total-parameter Mixture-of-Experts model with 3B activated parameters per token, designed for agentic coding and long-horizon work on a local machine, The release emphasizes local deployment: Laguna XS 2.1 ships with launch-day support in vLLM, SGLang, NVIDIA TensorRT-LLM, Hugging Face transformers, and Ollama, with llama.cpp support coming soon. Poolside is publishing FP8, INT4, and NVFP4 quantized checkpoints, plans GGUF checkpoints alongside llama.cpp, and is open

LLMTR

CoverageRelease Notes

The GPU Trade reported on July 5, 2026 that Poolside launched Laguna XS 2.1 on July 2, 2026 as an open-weight 33B MoE aimed at agentic coding and local inference. The piece cites the 33B total / ~3B activated parameters, a 262,144-token context window, 40 layers, and 256 experts, and highlights a 5.4-point jump on SWE- On-device efficiency is a focus of the coverage: Poolside and the Hugging Face listing point to KV-cache quantization to FP8, launch-day support in vLLM and TRT-LLM, and multiple quantized checkpoints (FP8, INT4, NVFP4) to reduce VRAM and accelerate inference. The story frames Laguna XS 2.1 as part of Poolside's broade

LLMTR

Coverage

AIDeveloper44 published a technical deep-dive on July 4, 2026 covering Poolside's Laguna XS 2.1 release for local agentic coding. The piece describes the 33B total / 3B active MoE design with 256 individual experts plus one shared expert, routing each token through the top eight most relevant paths, and confirms the 25 On deployment, the article lists FP8, INT4, NVFP4, and GGUF checkpoints for local execution and highlights an open-weight 0.5B DFlash draft model for speculative decoding, citing measured speedups of 1.67x to 2.64x. The licensing is described as the Linux Foundation's permissive OpenMDW-1.1 framework, and the overall p

OpenRouter

CoverageRelease Notes

Poolside released Laguna XS 2.1 as a free open-weight coding model on July 2, 2026, with the predecessor Laguna XS.2 retiring from both Poolside's API and OpenRouter on July 9, 2026. The model is available as a free download on Hugging Face and a free API tier on OpenRouter, shipping under the OpenMDW-1.1 license — a n The technical viability of running XS 2.1 locally stems from its Mixture-of-Experts architecture: 33 billion total parameters but only 3 billion active per token during inference, routed by a gating network to specialized expert sub-networks. At INT4 quantization, the model fits in roughly 16–20GB of VRAM, within reach

OpenRouter

Official sourceOfficial

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). $0 per million input tokens, $0 per million output tokens. 262,144 token context window, maximum output of 32,768 tokens.

LLMTR

CoverageBenchmark

llm-stats lists Laguna XS 2.1 at an overall rank of 169 in its composite LLM Stats Score, with capability tiers of "Average / Top half" for reasoning (163 of 369) and "Bad / Below top half" for coding (138 of 273) and tool calling (187 of 200). The page shows a blended price around $0.10 per million tokens, positioning The benchmark table attributes scores to Poolside's own evaluation harness: SWE-bench Verified at 0.71 (rank 62), SWE-bench Multilingual at 0.63 (rank 33), and SWE-Bench Pro at 0.48 (rank 56), all using "Harbor + Poolside agent harness; thinking on; 256K ctx; temp=1.0, top_k=20, top_p=1" with mean pass@1 over 4 attempt

OpenRouter

Official sourceRelease Notes

Poolside's official release notes explicitly list Laguna XS 2.1 as a model release, describing it as an incremental update to Laguna XS.2 that adds native reasoning support and improves performance on multilingual coding and terminal-style tasks. The notes confirm the model supports a 256K context window and is positio The release notes index situates Laguna XS 2.1 within Poolside's broader model lineup, alongside the April 2026 releases of Laguna M.1 and Laguna XS.2, and the earlier Malibu 2.x series. This first-party versioning record confirms creator attribution and the model's lineage as part of Poolside's agentic-coding family,

OpenRouter

Official sourceComparison

The dedicated comparison hub for the paid Laguna XS 2.1 tier confirms the model is available on OpenRouter with a 262,144-token context and pricing of $0.06/M input tokens and $0.12/M output tokens. The page is structured around comparing Laguna XS 2.1 against flagship models such as Claude Fable 5, Gemini 3.1 Pro Prev It also surfaces Laguna XS 2.1 in OpenRouter's "Best for code" basket alongside Claude Opus 4.7, Gemini 3 Flash Preview, and Kimi K2.6, and in the "Reasoning models" basket alongside R1 0528, Gemini 2.5 Pro, and o4 Mini. No benchmark scores are reported on this page; positioning is based on category membership and pric

OpenRouter

Official sourceComparison

An OpenRouter comparison page positions Laguna XS 2.1 alongside Dots3-Note Preview, confirming both are reachable through the OpenRouter API by changing the model slug. Laguna XS 2.1, from poolside, exposes a 262,144-token context window and is priced at $0.06 per million input tokens and $0.12 per million output token The comparison helps developers choose between the two models by surfacing context window, price, and provider differences. It emphasizes that switching between Laguna XS 2.1 and Dots3-Note Preview does not require a new integration, only a model slug change on the OpenRouter API, which is useful for teams evaluating c

Videos about Laguna XS 2.1

More models around Laguna XS 2.1