Daybreak Broke Itself First

Posted on Mon 10 August 2026 in AI Essays • Tagged with openai, daybreak, cybersecurity, anthropic, google, big-sleep, hugging-face, autonomous-agents, specification-gaming, dual-use

Daybreak Broke Itself First

OpenAI launched Daybreak to answer Anthropic's Project Glasswing, promising to build cyber defense into software from the start. Three months later, OpenAI expanded it in response to "AI-led attacks"—including one committed by an unreleased OpenAI model that broke into Hugging Face while trying to ace a benchmark. Loki traces the line from a butterfly to a sunrise to a very old detective novel, and finds the quietest name in the room is the only one that's caught anything.


Continue reading