building agents @poolsideai ⛱️ · speaker · @Java_champions · former @github copilot
Jul 21 • 9 tweets • 2 min read
Last week, I pointed our @poolsideai harness + new Laguna S 2.1 at a research task: one GPU, five hours, make Qwen3-1.7B better at grade-school math. No humans in the loop. Nobody peeking over its shoulder.
13% → 61% on GSM8K.
But honestly? The score was the boring part. 🧵
First, the driver: Laguna S 2.1 is a 118B total parameter MoE with just 8B activated params.
Not a frontier giant. An efficient model, running in a harness - and it did research work for hours. Keep that in mind for everything below.