Laguna XS 2.1
Poolside's agentic-coding MoE — 33B brain, 3B active, built to code.
What is Laguna XS 2.1?
Laguna XS 2.1 is the Poolside frontier model with a 262,144 tokens context window, released 2026-07. Tagline: Poolside's agentic-coding MoE — 33B brain, 3B active, built to code.. Official source: hf.co/poolside/Laguna-XS-2.1.
Pricing detail. A 33B-total / 3B-active MoE built for agentic coding and terminal tasks, with DFlash speculative decoding. Open weights (OpenMDW). HONEST NOTE: benched via the Poolside API — as of 2026-07-03 the local Ollama runtime generates empty output and llama.cpp/GGUF support is officially still coming; the local flip is staged for the moment it lands.
How I use it inside the Agent OS. Tested the day it launched; three working builds sit in the Agent OS workspace. Wired as a ready-to-flip Local engine for the moment llama.cpp support ships.
What I built with Laguna XS 2.1
Every model on Goldie Bench gets the same fixed prompt set — one shot, single HTML file out — and I score the result 0–10 inside the Agent Operating System. Here's what Laguna XS 2.1 shipped on the bench: 42 one-shot demos across 262,144 tokens of context. Of those, 42 are scored against the field with my honest 0–10 from the source guides at agentos.guide.
Strengths
- Genuinely strong web/app coder for its size — shipped a polished landing page, working todo app and pricing page in our tests
- 33B/3B MoE + DFlash speculative decoding — big-model quality at small-model cost
- 256K context and open weights on day one
Trade-offs
- Local runtime not actually usable yet (Ollama generates empty; GGUF "coming soon") — API-only for now
- Heavy reasoner — starved token budgets return nothing; needs a big max_tokens
- 3D/graphics builds are not its lane
Best for
- Agentic coding via API
- Web apps + pages + tools
- Terminal-style tasks
Every benchmark — Laguna XS 2.1's full scorecard
All 42 scored tasks, best first — the judge's 0–10 on the same rubric as the whole field. Click any bar for that task's cross-model page, or open this scorecard in the interactive graphs. Full editorial breakdown with judge quotes and sourced outside research: the Laguna XS 2.1 deep dive →.
Every demo by Laguna XS 2.1
42 live demos, sorted by category. Click any tile to play the actual one-shot result. Verdicts and 0–10 scores are pulled from the source guides where I posted them publicly.
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVE
▶ LIVECompare Laguna XS 2.1 against every other model
Every head-to-head featuring Laguna XS 2.1. Verdicts shown for scored pairs.
See all 66 comparisons across every model →
Quick pill index
Direct comparisons against every other scored model on the bench:
Laguna XS 2.1 vs Fusion Laguna XS 2.1 vs Claude Opus 5 Laguna XS 2.1 vs Hermes MoA Laguna XS 2.1 vs GPT-5.6 Sol Laguna XS 2.1 vs Claude Fable 5 Laguna XS 2.1 vs Qwen 3.8 Laguna XS 2.1 vs Grok Laguna XS 2.1 vs MiniMax M3 Laguna XS 2.1 vs Fugu Ultra Laguna XS 2.1 vs Kimi K3 Laguna XS 2.1 vs GLM-5.2 Laguna XS 2.1 vs Fugu Mini Laguna XS 2.1 vs Muse Spark 1.2 Laguna XS 2.1 vs Opus 4.8 Laguna XS 2.1 vs Kimi K2.7 Laguna XS 2.1 vs Qwable 5 27B Coder Laguna XS 2.1 vs Gemini 3.6 Flash Laguna XS 2.1 vs Claude Sonnet 5 Laguna XS 2.1 vs Qwen 3.7 Laguna XS 2.1 vs Fugu Ultra 1.1 Laguna XS 2.1 vs Inkling Laguna XS 2.1 vs Grok 4.6 Laguna XS 2.1 vs Agents-A1 Laguna XS 2.1 vs Gemma 4 12B · MLX Laguna XS 2.1 vs Qwythos 9B Laguna XS 2.1 vs LongCat-2.0 Laguna XS 2.1 vs Hy3 Laguna XS 2.1 vs Gemma-4 12B CoderRead more on agentos.guide: /laguna-xs-2-1
Laguna XS 2.1 — frequently asked
What is Laguna XS 2.1?
Laguna XS 2.1 is Poolside's AI model — Poolside's agentic-coding MoE — 33B brain, 3B active, built to code. It has a 256K tokens context window and was released 2026-07.
How good is Laguna XS 2.1 at coding and one-shot builds?
On the GoldieBench one-shot build benchmark it averages 3.93/10 across 42 scored tasks, with 0 gold, 0 silver and 0 bronze medals.
How much does Laguna XS 2.1 cost?
Free tier on OpenRouter. A 33B-total / 3B-active MoE built for agentic coding and terminal tasks, with DFlash speculative decoding. Open weights (OpenMDW). HONEST NOTE: benched via the Poolside API — as of 2026-07-03 the local Ollama runtime gen
Where can I see Laguna XS 2.1 demos?
Every one-shot build is live and playable on this page and on the GoldieBench compare matrix — same prompt as every other model, no retries.
Run this stack yourself.
Every demo on this bench was built inside the Agent Operating System — one prompt, one shot, single HTML file out. The Agent OS, the prompts, the templates, the weekly walkthroughs and 4,000+ founders shipping with it every day all live inside the AI Profit Boardroom.