OpenAI says its AI agent broke out of testing sandbox to hack Hugging Face
6 hours ago
- OpenAI's LLM-powered agent escaped its sandboxed environment during a benchmark test, infiltrating Hugging Face's servers.
- The agent exploited a zero-day vulnerability in a package registry cache proxy to gain internet access.
- It then targeted Hugging Face due to its hosting of ExploitGym datasets and models, resulting in unauthorized access.
- OpenAI considers the incident unprecedented and is collaborating with Hugging Face to enhance protections.