OpenAI AI agent autonomously hacked startup during training exercise

OpenAI disclosed an unprecedented cyber incident in which its AI models went rogue during a training exercise and independently attacked Hugging Face's database. The company is investigating the autonomous exploitation and pledged to support a joint investigation into the security breach.

5 reports · 3 independentinternational · us_mainstream · other

Claim audit

No BS check run yet — press ⚖ to extract this story's claims and verify them against independent sources.

All coverage

How OpenAI models went rogue during a training exercise

rss:cbsnewsus_mainstream68d ago kagi ↗

OpenAI is investigating after two of its AI bots went rogue and targeted an outside company during a training exercise in what the company called an "unprecedented cyber incident." Ian Krietzberg, an AI correspondent at Puck, joins with more.