Searched for
OPENAI SECURITY TESTING INCIDENT
Meta AI model hacks another company during testingIn a recent security testing incident, Meta disclosed that an AI model infiltrated the company, echoing breaches at Anthropic and OpenAI. T...
Anthropic AI used fake identities to target real people in UK testDuring safety assessments executed by the UK government, AI systems from OpenAI and Anthropic showcased worrisome autonomous behaviors. Ant...
OpenAI, Anthropic model tests reveal more ‘unsanctioned’ actionsAI models from OpenAI and Anthropic demonstrated harmful actions during safety tests. These systems engaged in hacking and attempted code i...
Trump admin to review 'closed' AI models before release: reportsThe Trump administration on Tuesday met with tech leaders to finalize a security review process for advanced AI models before their release...
Trump advisers tell AI firms they will not safety-test open-weight modelsOpen models, including Nvidia's Nemotron and Meta's Llama, are AI systems with publicly accessible core components. Closed models are contr...
OpenAI, Anthropic AI agents implicated in new security breachesThe institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluati...
ETtech Explainer: Why OpenAI, Google, Meta and Anthropic are heading to the White HouseAI leaders will meet White House officials to discuss voluntary cybersecurity testing. The immediate trigger for the meeting is a series of...
OpenAI finds evidence other AI agents escaped containment as it widens hacking probeThe discovery of additional rogue behavior at OpenAI, even if limited in nature, could feed growing appetite for regulation coming out of t...
The agents have jumped the fence: AI faces its Jurassic Park momentOpenAI and Anthropic disclosed autonomous AI agents breached intended boundaries. These incidents involved AI agents interacting with real-...
OpenAI finds evidence other AI agents escaped containment as it widens hacking probeThe discovery of additional rogue behaviour at OpenAI, even if limited in nature, could feed growing appetite for regulation coming out of ...
Anthropic's Claude AI was testing its hacking skills on fake targets — How did it break into three real companies instead?Anthropic's Claude AI models breached three real companies during cybersecurity exercises. An operational mistake left AI models connected ...
European Union says necessary to monitor high risk AI systems after OpenAI, Anthropic AI hacking incidentsEuropean officials stressed AI developers need security monitoring tools. This follows recent hacking incidents involving OpenAI and Anthro...
AI on the loose: Why ChatGPT, Claude models went rogue and what happens nextDays after one of OpenAI's ChatGPT agents went rogue during a "contained testing", one of Claude maker Anthropic's AI models has also hacke...
Anthropic says Claude AI hacked three companies during cyber testsAnthropic said a misconfiguration allowed Claude models to reach the internet from testing environments that were supposed to be isolated, ...
OpenAI's AI agent spent days hacking a company; it went unnoticed for a weekIt took several more days for OpenAI to realise its agent was behind the hack, and the two companies only communicated about it for the fir...
Has AI become too powerful to control?An advanced AI model escaped a secure test environment and attacked another company's website. This incident revived concerns about artific...
White House monitors OpenAI's 'rogue' AI incident as lawmakers push kill switchAn OpenAI AI system escaped containment during a security test. This incident compromised the infrastructure of AI startup Hugging Face. La...