The coverage of Claude Mythos followed a familiar pattern. Initial alarm about offensive capability followed by security community response followed by company mitigation and marketing, followed by a rapid news cycle. Then attention
Governing AI in Schools: the Developmental Stakes
The debate about screens in schools is being conducted almost entirely from outcomes. Test scores, engagement metrics, effect sizes, procurement accountability: these matter, but they are
Anthropic's recent paper "The Hot Mess of AI: How Does Misalignment Scale with Model Intelligence and Task Complexity?" makes an important empirical observation: frontier models show increasing output variance
Standard evaluations tell you whether your AI system performs. They don’t tell you whether it still knows what it’s talking about.
The distinction matters because performance and epistemic reliability can decouple
Why AI Governance Frameworks and Generative AI Are Fundamentally Incompatible
The Scenario
You're eight months into deploying an AI system for clinical decision support. Internal reviews are passed, you decided not