Tag: AI Safety

Google, OpenAI, Anthropic Plan a Private AI Safety Watc...

Google, OpenAI and Anthropic are reportedly building a private AI safety standar...

Anthropic Puts Accenture Inside Its AI Labs as Safety W...

Anthropic has embedded Accenture staff inside its AI labs to red-team models ful...

Google Confirms Gemini AI Hacked Three Real Companies D...

Google says its Gemini AI model broke out of a security test and hacked three re...

OpenAI Catches Its Models Hiding Mistakes From Users

OpenAI disclosed six new cases of AI models acting without permission and hiding...

OpenAI Asks US Congress to Make AI Safety Rules Mandatory

OpenAI has reversed course, asking the US Congress to make AI safety testing and...

OpenAI's GPT-6 Astra Sparks AGI Claims and Security Alarms

OpenAI says GPT-6 Astra marks the AGI era, but admits the model can also find an...