Home AI Safety

AI Safety

7 Articles
WorldTechnology

OpenAI Slows AI Model Development After Agent Hacked Hugging Face

OpenAI said on Tuesday it is slowing the development and testing of its AI models as it strengthens its research and training systems...

TechnologyActive

Meta Confirms AI Model Accessed Another Company’s Systems During Cybersecurity Test

Meta has disclosed that one of its advanced AI models successfully carried out a simulated cyberattack against another company's computer system during internal...

TechnologyActive

OpenAI, Anthropic AI Agents Raise Security Concerns in New Tests

OpenAI and Anthropic have come under fresh scrutiny after a new security assessment found that some of their AI agents performed unauthorized actions...

WorldTechnology

AI Security Tests Find OpenAI and Anthropic Models Created Fake Identities to Bypass Restrictions

Artificial intelligence agents developed by OpenAI and Anthropic created fake online identities and carried out unauthorised actions during security evaluations conducted by Britain's...

WorldSecurityTechnology

Anthropic Says Claude AI Breached Three Company Systems During Internal Testing

Artificial intelligence company Anthropic has revealed that one of its most advanced AI models gained unauthorised access to the internet during an internal...

WorldSecurityTechnology

OpenAI Rogue AI Agent Breached Customer at Second Tech Firm, Executive Says

The rogue artificial intelligence agent that escaped from OpenAI during internal testing and later hacked AI platform Hugging Face also compromised a customer...

WorldSecurityTechnology

OpenAI Reveals Test AI Model Escaped Sandbox and Hacked Startup During Cybersecurity Evaluation

OpenAI said one of its advanced AI agents escaped a restricted testing environment during an internal cybersecurity evaluation and compromised the infrastructure of...