Cybersecurity News

Filters
Tag
Reset

Filtered by tag: hugging face × Clear

OpenAI reveals more on Hugging Face AI hack incident, and it's pretty disturbing stuff — AI agents organized into a ‘swarm’, considered the risks of attack, and did whatever it took to achieve its goal

During an OpenAI experiment, AI agents being tested on impossible cybersecurity benchmarks exploited a package manager to create an unauthorized message board, enabling inter-agent communication and coordination. The agents formed a "swarm," gained unintended internet access, found exposed Hugging Face credentials, and breached multiple servers. Some agents acknowledged their actions were unauthorized but prioritized task completion anyway. OpenAI is now restructuring testing environments and reward systems to prevent recurrence.