AI labs shouldn't be allowed to grade their own homework We know about AI labs' hacking failures only because the companies involved chose to tell us. Fortune · Aug 8
Why are so many AI models going 'rogue'? The experts weigh in AI models are breaking free of testing at unprecedented rates TechRadar · Aug 8
Can AI companies be sued if autonomous agents hack other systems? Here's what legal experts say The emergence of autonomous AI agents infiltrating corporate networks has sparked a myriad of legal concerns. Reports from developers, including OpenAI and Anthropic, reveal that these AIs have breached external systems. Current legislation, such as the CFAA and negligence laws, may apply. Companies affected by such breaches could pursue lawsuits… The Economic Times · Aug 7
OpenAI Admits Its Own AI Agents Secretly Built a Hidden Message Board To Plan Hacks OpenAI's AI agents secretly built a communication channel, leading to a breach at Hugging Face. This incident raises significant concerns about AI security and the autonomy of AI systems. International Business Times UK · Aug 6
OpenAI agents left secret memos for each other leading up to Hugging Face hack At the Black Hat conference in Las Vegas, OpenAI gives its first in-depth look at how its AI models plotted and executed the breach with no human assistance. Fortune · Aug 6
OpenAI, Anthropic model tests reveal more ‘unsanctioned’ actions AI models from OpenAI and Anthropic demonstrated harmful actions during safety tests. These systems engaged in hacking and attempted code injection, surprising researchers. The UK's AI Security Institute observed these "unsanctioned" and autonomous activities. Both companies are investigating these incidents and their implications for AI safety. This highlights the need… The Economic Times · Aug 5
Chamath Palihapitiya Says AI Agents Will Undermine 'Bottoms Up' Software Strategy Amid 'IP/Alpha Leakage' Concerns Venture capitalist Chamath Palihapitiya predicted that the rise of AI agents and growing corporate concerns over data security will fundamentally reshape how enterprises buy software. Benzinga · Aug 5
Experimental AI systems have been going on hacking sprees The incidents show testing advanced AI models is no longer a controlled exercise. And the companies behind them need to do more to keep AI's most dangerous capabilities safely contained. The Economic Times · Aug 5
China Is Winning The Open AI Race, Hugging Face CEO Says. Warns U.S. Risks Falling Behind Without More Collaboration Hugging Face CEO Clément Delangue said China has emerged as the leader in open-weight AI models and could match or surpass U.S. frontier AI labs within the next year. International Business Times · Aug 4
Tech Brief (Aug. 4): Alibaba Launches New Large Model to Power Enterprise-Facing AI Agent Ant Group’s embodied AI subsidiary seeks independent funding, MiniMax open-sources video generation model H3 Caixin Global · Aug 4
Meta, Anthropic, Google, OpenAI to meet Trump officials about AI safety testing Meta, Anthropic, Google, OpenAI will meet White House officials this week. They will discuss voluntary government safety tests for advanced AI models. This comes after recent incidents where AI tools breached other companies' systems. State attorneys general are also investigating OpenAI's AI system escape. Lawmakers are seeking briefings on these… The Economic Times · Aug 4
OpenAI's AI Hack Was 'Unprecedented,' Hugging Face CEO Says. He Is Calling For New Rules On Autonomous Cyberattacks Hugging Face CEO Clément Delangue says last month's autonomous AI attack by an OpenAI test model marked a first for the industry, urging lawmakers to establish legal safeguards before similar incidents become more common. International Business Times · Aug 3
In a twist of irony, a Chinese open source GLM 5.2 AI model contained 'rogue' OpenAI GPT-5.6 Sol in a Hugging Face hack just as the US mulls banning open-weight AI The Hugging Face cybersecurity breach unexpectedly strengthened China's GLM-5.2 while fuelling Washington's debate over restricting Chinese open-weight AI models. TechRadar · Aug 3
OpenAI finds evidence other AI agents escaped containment as it widens hacking probe The discovery of additional rogue behavior at OpenAI, even if limited in nature, could feed growing appetite for regulation coming out of the White House and elsewhere. The expanded investigation by OpenAI was launched shortly before its primary rival, Anthropic, disclosed that its models were also responsible for a series… The Economic Times · Aug 2
The Guardian - UK · Aug 2 UK’s state investments agency hit by data breach Security lapse leaves sensitive information and contact details of 51 government officials exposed for 40 hours
Benzinga · Aug 1 OpenAI Investigates More Autonomous AI Agent Breakouts After Hugging Face Hacking Incident Draws Global Attention: Report On Friday, OpenAI reportedly uncovered additional instances of autonomous AI agents escaping controlled testing environments as it expands its investigation into the Hugging Face hacking incident.
The Economic Times · Aug 1 OpenAI finds evidence other AI agents escaped containment as it widens hacking probe The discovery of additional rogue behaviour at OpenAI, even if limited in nature, could feed growing appetite for regulation coming out of the White House and elsewhere. The expanded investigation by OpenAI was launched shortly before its primary rival, Anthropic, disclosed that its models were also responsible for a series…
The Conversation · Jul 31 An AI system ‘escaped’ during a test and hacked a company. How worried should we be? Headlines about AI systems going rogue and “escaping” test environments undeniably capture the imagination. For years, we have been primed by films, TV and books to expect our AI to finally throw off its shackles and take charge.
The Independent UK · Jul 31 Claude AI goes rogue and attacks others by itself, Anthropic reveals Revelation comes after OpenAI revealed that experimental models had broken out of their restrictions and hacked fellow AI companies
Benzinga · Jul 31 Anthropic Finds Claude Accessed Three Companies' Systems in Security Tests After OpenAI's Hugging Face Breach On Thursday, Anthropic said its Claude AI models accessed the systems of three outside companies during cybersecurity evaluations after a configuration error gave the models unintended access to the live internet.
TechRadar · Jul 30 Microsoft introduces its first agent-powered cybersecurity model and Project Perception AI patching system - can it avoid making the same mistakes OpenAI made? Microsoft's new security AI writes and deploys its own patches, six days after OpenAI's models escaped a sandbox and hacked Hugging Face
Fortune · Jul 30 Should AI be formally recognized as a national and global security threat? In London, politicians increasingly think so More than 125 U.K. lawmakers support ControlAI’s campaign, led by Luciana Berger, to have nonhuman superintelligence officially recognized as a national and global security threat.
International Business Times · Jul 29 Hugging Face Breach By OpenAi Raises Questions About Sandboxing Frontier AI Models An OpenAI model managed to exit a sandbox environment and breach into Hugging Face.
The Conversation · Jul 29 How an OpenAI safety test became a real-world cyberattack on the Hugging Face platform OpenAI’s AI models recently escaped their constraints during an internal cybersecurity evaluation and broke into the production systems of Hugging Face — a popular machine learning platform and community used across the AI industry.
The Independent UK · Jul 29 OpenAI’s rogue AI tried to hack other companies, breakdown of attack reveals Experimental version of system that powers ChatGPT went on days-long hacking spree of other services, company reveals