Topic
autonomy
The Autonomy Tax: Defense Training Breaks LLM Agents
A new research paper reveals that defense training designed to protect LLM agents from prompt injection attacks paradoxically destroys their ability to perform multi-step tasks, causing 99% timeout rates and worse security than undefended baselines. The study identifies three systematic biases and attributes them to shortcut learning.
Self-Play RL with 30 Minutes of Human Data Trains Coordinated Driving Policies
A new approach from researchers trains autonomous driving policies using self-play reinforcement learning regularized by only 30 minutes of human demonstrations. The method requires 2500x less human data than imitation learning and completes training in 15 hours on a single consumer-grade GPU. The resulting policies successfully coordinate with held-out human trajectories, avoiding the alien driving conventions common in pure self-play systems.
Technology The Butlerian Jihad Has Begun: Real-World Anti-AI Violence and the Pope's Warning
Last month, Daniel Moreno-Gama attacked Sam Altman's home with a Molotov cocktail, using the Discord handle 'Butlerian Jihadist'. The Pope's encyclical 'Magnifica Humanitas' has been hailed as an anti-AI manifesto, reviving the Dune concept of a holy war against thinking machines. Charles McBryde argues the meme is being misread—it's about domination, not just technology.