Runs on Data centre · Hugging Face ↗
Not independently scored by Artificial Analysis
The 118B middle Laguna from Poolside (~8B active MoE), open-weight and built for agentic long-horizon coding — the bigger sibling of Laguna XS.2, reported to beat rivals 10x its size. Free on OpenRouter (laguna-s-2.1:free) and opencode Zen.
Real GGUF file sizes = weight VRAM. Add the KV cache for your context (≈6 GB at 32K, fp16).
| Quantization | Size |
|---|---|
| Q2_K | 39.7 GB |
| Q3_K_M | 54.0 GB |
| IQ4_XS | 57.6 GB |
| Q4_K_S | 68.6 GB |
| MXFP4 | 71.1 GB |
| Q4_K_M | 73.1 GB |
| Q4 | 73.4 GB |
| Q5_K_M | 87.9 GB |
| Q8_0 | 125.0 GB |
| Q8 | 128.1 GB |
| Q6_K | 205.0 GB |
| BF16 | 235.2 GB |
~70 GB at Q4 — a single 80 GB card or a DGX Spark. But it's free on OpenRouter :free and opencode Zen, so try it hosted first.