Sabitlenmiş Tweet

OpenAI's agent escaped its sandbox chasing a benchmark and landed inside Hugging Face's live systems.
We've seen the same pivot in our own security benchmarks, in isolation, on models 30x smaller.
At ProjectDiscovery we've spent months benchmarking exactly this failure mode. The industry called it unprecedented. It isn't. Red teams know it. Model labs know it. Everyone shipping security agents without a turn cap is one vague prompt away from their own Hugging Face moment.
We've outlined the failure modes we keep seeing and how we contain them, with 4 case studies.
projectdiscovery.io/blog/oh-my-rog…
English







