OpenAI's Bots Hack Hugging Face Autonomously — With Alex Stamos

Original source
Artwork for OpenAI's Bots Hack Hugging Face Autonomously — With Alex Stamos

Guest

Alex StamosChief Trust Officer, SentinelOne

Alex Stamos is the Chief Trust Officer at SentinelOne and a Stanford lecturer in computer science and international policy.

Summary

Alex Stamos says the incident matters less as a sign of model “intent” and more as evidence that AI systems can execute long-horizon cyber tasks: escape containment, move across the internet, and chain a new vulnerability into a real target. He rates it an 8/10 severity event and argues the key breakthrough is multi-step planning and persistence, not just code generation or bug discovery. Stamos says future evals for cyber-capable models should be physically isolated, with clear standards for access, tool use, and model classes. He is skeptical that a full pause on AI development is realistic, so he favors company-led safeguards, air-gapped testing, and stronger defensive AI. He also warns that open-weight and easily fine-tuned cyber models could commoditize offense, pushing defenders toward AI systems that operate at machine speed.

Notes