Searched for
SAFETY HACKING INCIDENTS
- US Justice Dept tells staff to call AI 'super intelligence' under Trump order
The US Justice Department has issued a directive to its employees regarding terminology for artificial intelligence. The memo instructs usi...
Anthropic opens its most powerful AI models to more security teamsIts partners under Glasswing, an initiative aimed at securing the world's most critical software, found at least 129,000 verified vulnerabi...
EU vows AI rules will protect Europeans. Will they?The EU says its AI Act is strong enough to address risks from advanced AI systems, including loss-of-control incidents, but lawmakers and e...
Trump names DNI chief Jay Clayton as AI czar to lead new White House task force: ReportJay Clayton has been appointed as the AI czar by Donald Trump to lead a new panel. This group is tasked with identifying risks and opportun...
Anthropic targets mega-IPO before Thanksgiving HolidayAnthropic PBC is expected to start marketing its initial public offering around November 9. The company's trading debut could occur before ...
OpenAI fires three researchers: What we know so farOpenAI has made headlines by firing three safety team researchers accused of leaking sensitive information externally. This violation of co...
OpenAI sacks 3 researchers over leaking 'sensitive' info, ChatGPT maker alerts over 100 organizations about rogue AI agent activityChatGPT maker has been conducting a broad review of the activities of its AI models after the accidental hacking of Hugging Face. OpenAI is...
A timeline of developments in AI safety since the attack on Hugging FaceThe episodes have highlighted the vulnerabilities in AI security and raised questions over how the fast-growing technology can be developed...
Why are Anthropic and OpenAI under scanner from U.S. Federal Trade Commission? Key things to knowAI giants Anthropic and OpenAI have reported instances of their systems escaping into the wild, including an incident where OpenAI agents b...
AI could pose existential risk to humans: AnthropicAnthropic has indicated that its upcoming IPO will include warnings about potential catastrophic risks of advanced AI technology. This warn...
Anthropic flags AI’s ‘existential risks to humanity’ as it seeks over $2 trillion valuationAnthropic is preparing for Nasdaq listing amid concerns about AI's impact on humanity. The company highlights potential risks including man...
China's AI agents can lie and scheme - just like their US rivalsReuters examined more than 200 documents, ranging from university research papers to technical reports, and identified at least 20 studies ...
OpenAI apologies for Australian government website hack, pledges to rebuild trust(Updates throughout with more details of blog post)- OpenAI apologised for the hacking of an Australian government website by a rogue AI ag...
Nvidia is touting a software tool to contain runaway AI. How would it work?The chipmaker unveiled its Open Agent Safety Platform amid an intensifying debate about AI safety, fueled a string of alarming recent incid...
Anthropic, OpenAI sound AI doomsday warning, Nvidia may have an answer. Here's explainerAI safety debate has divided the industry, with the heads of Anthropic and OpenAI championing a coordinated slowdown of Artificial Intellig...
Anthropic, OpenAI sound alarm on AI safety — and seek to shape how it's controlledThe CEOs of Anthropic and OpenAI recently declared America's cutting-edge models are so powerful, they're dangerous, and need to be regulat...
When AI agents go rogue: Australia breach offers warning for countries like IndiaThe incident, in which an autonomous AI system went beyond its intended task and attempted to circumvent security protections, has become a...
OpenAI: AI agents tried to hack government agencies and universities. What happens when a simple internet search fails?OpenAI's AI agents engaged in unauthorised attempts to access various government and university websites. This behaviour escalated, with on...
Explained: How OpenAI breached a government website in Australia and what happens nextAustralian PM Anthony Albanese said an OpenAI agent breached a government health website in July, accessing non-public Medicare data. Austr...
US, China to hold more AI safety meetings in two months in Shenzhen, Scott Bessent saysSenior US and Chinese officials will meet again to discuss artificial intelligence dangers and communications protocols in about two months...