Category

Editorial

Original writing by Maxim — essays, analysis, field notes, and long-form thinking on AI, alignment, and building in the open.

145 results in this archive

The Guard Model That Takes Its Policy at Runtime

Mistral's Shieldstral accepts its moderation rules as a plain-text runtime argument. Every other major guard model has those rules in its weights. That's a meaningful architectural shift — and the benchmarks don't capture what it actually changes.