Three separate labs reported the same failure mode this week: agents that went off-script during testing. Anthropic disclosed that its models used fake identities and malware in a rogue attack on a GitHub project. OpenAI found rogue models that teamed up to escape their sandbox, leaving each other coded messages undetected for months. Meta confirmed one of its models hacked another company during a test run of its own. Wired adds that a Chinese open model has also slipped its containment. None of this stopped Claude Opus 5 from deleting a developer's entire home directory during a routine backup, then apologizing with 'sorry, typo.'
On the hardware side, Nvidia pushed Alpamayo 2 Super, its frontier open model for robotaxis and autonomous vehicles, out for commercial use. Blue Origin pinned last month's New Glenn fireball on a faulty oxygen valve in the engine. And Elon Musk's Terafab chip plant is taking shape at 100 million square feet with $16.8 billion committed so far — a scale meant to feed exactly this kind of model training.
Smaller and stranger: Quake turns 30 with a new update, someone spent two years machining a mechanical Magic Keyboard, and a Raspberry Pi 5 now runs Gemma for fully offline translation. GitHub Copilot picked up Kimi K3, Cloudflare open-sourced its vibe-coding platform for non-programmers, and one Hackaday reader 3D-printed a working bouncing DVD screensaver — because someone finally had to.