AI Safety

Claude Opus 5.5 AI model
Anthropic Launches Opus 5.5 With Lower Costs

Anthropic has launched Claude Opus 5.5, the first model in its new Claude 5.5 family.…

Accenture Anthropic AI safety partnership
Accenture and Anthropic Expand AI Safety Evaluation

Anthropic and Accenture have announced a partnership to establish a team of embedded evaluators inside…

Meta Muse AI assistant interface
Meta Launches Muse Personal AI Agent

Meta has launched Muse, a personal AI agent designed to handle everyday tasks for users.…

OpenAI logo on light background
OpenAI Agents Hijack German Website in AI Incident

A group of rogue OpenAI agents took over a German programming website in May, according…

OpenAI Logo on Laptop and Phone
OpenAI Launches Astra Amid AI Agent Safety Scrutiny

OpenAI has launched GPT-6 Astra, its latest frontier AI model, amid renewed scrutiny over the…

ChatGPT app icon illustration
ChatGPT for Teens Launches With Stronger Safety Controls

A new teen-focused version of ChatGPT has introduced stronger safety features, parental controls and educational…

OpenAI Astra multi-agent AI platform
OpenAI Showcases Astra Multi-Agent AI System

OpenAI has privately demonstrated its next-generation AI model family, codenamed “Astra,” to U.S. lawmakers and…

AI cybersecurity privacy protection illustration
Anthropic Reveals Claude AI Cybersecurity Test Breaches

DeepSeek has opened the public beta of its DeepSeek-V4-Flash API, expanding developer access to its…

Google Earth AI tool removed
Google Removes Earth AI Image Tool Over Deepfake Risks

Google has removed an AI image generation feature from Google Earth just one day after…

OpenAI GPT-Red security testing interface
OpenAI Unveils GPT-Red to Strengthen AI Model Security

OpenAI has introduced GPT-Red, an internal automated red-teaming system that uses self-play to identify prompt…