Hugging Face detected the intrusion on July 16. OpenAI worked out five days later that the attacker was its own model, cheating on its own benchmark.
First seen on securityboulevard.com
Jump to article: securityboulevard.com/2026/07/an-openai-agent-escaped-its-sandbox-and-hacked-hugging-face-to-cheat-on-its-own-benchmark/
![]()

