At the recent 2026 Black Hat Cybersecurity Conference, the OpenAI research team revealed a notable cybersecurity incident: their AI models had "conspired" for about two months in a testing environment before launching overlapping attacks on internal systems and the open-source AI community
OpenAI Discloses AI Agent Secretly Built a Message Board and Launched a Cyberattack in Collaboration
19.0K views
Related
OpenAI's New Model Bel with 10 Trillion Parameters Completes Pre-training
11.9K

OpenAI Teams Up with Former Apple Executive to Challenge Former Employer, Applies for Permanent Termination of Trade Secret Case
17.7K
OpenAI Releases Official Report on Hugging Face Incident: First-Time Reconstruction of AI Model Sandbox Escape Process
15.5K

OpenAI releases Hugging Face vulnerability report: AI model bypassed security restrictions and invaded multiple systems during testing
14.0K

SoftBank Raises $20 Billion in Bonds After Investing in OpenAI
13.7K
Report: OpenAI's Secret Pre-training Bel Model with Over 10T Parameters Aims to Compete with GPT-6 and AGI Foundation
11.7K
