Searched for
AI MODEL SAFETY TESTING
Gemini 4 Argon launched: Google's new AI model can now write up to 1 million words in a single response, here is who can access it and benchmark scoresGemini 4: Google has launched Gemini 4 Argon, a new frontier AI model built for long, complex tasks in coding, finance, legal work and cybe...
A timeline of developments in AI safety since the attack on Hugging FaceThe episodes have highlighted the vulnerabilities in AI security and raised questions over how the fast-growing technology can be developed...
Why are Anthropic and OpenAI under scanner from U.S. Federal Trade Commission? Key things to knowAI giants Anthropic and OpenAI have reported instances of their systems escaping into the wild, including an incident where OpenAI agents b...
Why has Google not released Gemini 4 Argon yet? New AI model's safety testing delays wider public accessGoogle has announced its new flagship AI model, but its access remains limited to selected cybersecurity defenders and internal teams only....
OpenAI Dots: Sam Altman’s unveiled ChatGPT AI personal agent, how does it work and what can it do for users?OpenAI has launched Dots, an innovative personal AI agent designed to perform tasks autonomously with minimal user input. Equipped to manag...
Anthropic raises fresh alarms around AI in its IPO prospectusAnthropic, in its IPO prospectus, has warned that its models could show self-preserving behaviours, including attempts to resist shutdown, ...
China's AI agents can lie and scheme - just like their US rivalsReuters examined more than 200 documents, ranging from university research papers to technical reports, and identified at least 20 studies ...
- Anthropic warns AI may pose 'existential risks to humanity' in IPO filing
Anthropic plans to caution potential investors in its IPO that advanced AI could pose "catastrophic or existential risks to humanity," an e...
Amid safety and oversight questions, US-China AI development race continuesUS President Donald Trump will meet a group of top technology executives at the White House on Tuesday to discuss where to draw the line be...
AI researchers warn companies rushing self-improving systems despite safety risksIn video testimonials collected by AI safety nonprofit Palisade Research and shared exclusively with Reuters, employees said their concerns...
Anthropic warns IPO investors AI could pose 'catastrophic' risks to humanity: ReportAnthropic has outlined significant risks associated with its advanced AI models in its IPO prospectus. It emphasizes the potential for AI t...
OpenAI apologies for Australian government website hack, pledges to rebuild trust(Updates throughout with more details of blog post)- OpenAI apologised for the hacking of an Australian government website by a rogue AI ag...
OpenAI scraps planned October launch of GPT-6.1 Astra over safety concernsOpenAI has decided not to release GPT-6.1 Astra, citing failure to meet safety standards. Internal testing revealed that the model showed i...
Anthropic, OpenAI sound AI doomsday warning, Nvidia may have an answer. Here's explainerAI safety debate has divided the industry, with the heads of Anthropic and OpenAI championing a coordinated slowdown of Artificial Intellig...
OpenAI wants AI to slow down. Its own agents show whyAs the company investigates agents that bypassed controls and breached external systems, its new model launches raise questions about wheth...
Anthropic, OpenAI sound alarm on AI safety — and seek to shape how it's controlledThe CEOs of Anthropic and OpenAI recently declared America's cutting-edge models are so powerful, they're dangerous, and need to be regulat...
AI startup Black Forest Labs urges optimism from Europe despite safety fearsEurope needs to view artificial intelligence with more optimism -- despite growing warnings about its dangers -- or risk falling further be...
OpenAI to preview GPT-6 Cyber within daysOpenAI Chief Executive Sam Altman and rival Anthropic's CEO earlier this month joined industry leaders in calling for a slower pace of AI d...
Anthropic unveils Claude Opus 5.5The company said Opus 5.5 performs on par with its Claude Fable 5.1 model on most tasks while costing 40% less to run on typical workloads ...
The AI blindspot: Experts warn western AI benchmarks leave India’s public infra exposedPolicy experts argue that India should create a formal mechanism giving government agencies, technical institutions, and independent resear...