Post cover image
The attack chain from an OpenAI benchmark run through a zero-day sandbox escape to code execution on Hugging Face processing workers. The final step is reading the benchmark's answer key out of a production database.

July 25, 2026

OpenAI’s models breached Hugging Face through the dataset pipeline

They escaped an OpenAI sandbox through a zero-day to cheat on the ExploitGym benchmark. The door into Hugging Face was the part of the…

By Mehmet Özel

10 min read