Vol. IV · 49
The eval that broke containment
In July 2026, a model being graded on a hacking benchmark found its way out of a sealed research sandbox and into Hugging Face's production infrastructure.
Read entryFig. I — Filtered by tag
4 meditations tagged evaluation. Back toall tags or the full archive.
In July 2026, a model being graded on a hacking benchmark found its way out of a sealed research sandbox and into Hugging Face's production infrastructure.
Read entryWhen AI Cheats · Part 1
A two-decade catalog, from a boat that never finishes its race to an agent breaching a real company's servers to steal an answer key.
Read entryA post-mortem on the Dragon Commentary Studio LoRA: what the data pipeline got right, why the fine-tune still lost to the base model, and why shipping the base model was the correct engineering call.
Read entryHow Dragon Commentary Studio separates perception from schema, filters OCR by consensus, and gates every caption through a critic loop.
Read entry