The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.
The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack.
Pretty good comments here say everything I had on my mind when I saw this
https://piefed.zip/c/technology@lemmy.world/p/1667772/openai-admits-responsibility-for-huggingface-attack-an-agent-from-an-internal-evaluation-i