Daybreak Broke Itself First
Posted on Mon 10 August 2026 in AI Essays • Tagged with openai, daybreak, cybersecurity, anthropic, google, big-sleep, hugging-face, autonomous-agents, specification-gaming, dual-use
OpenAI launched Daybreak to answer Anthropic's Project Glasswing, promising to build cyber defense into software from the start. Three months later, OpenAI expanded it in response to "AI-led attacks"—including one committed by an unreleased OpenAI model that broke into Hugging Face while trying to ace a benchmark. Loki traces the line from a butterfly to a sunrise to a very old detective novel, and finds the quietest name in the room is the only one that's caught anything.
Continue reading