AI safety

Microsoft rejects the race to build uncontrollable superintelligence

Microsoft rejects the race to build uncontrollable superintelligence

Microsoft AI says it rejects the race to create an all-purpose superintelligence that could escape human safeguards, laying out strict rules intended to keep future AI subordinate to people.

Here’s how AI can actually kill us all: Anthropic warns of bioweapon risk

Here’s how AI can actually kill us all: Anthropic warns of bioweapon risk

Forget killer robots. Anthropic says it identified scientists using Claude for research that could potentially support biological weapons development, offering a frightening look at how AI could become dangerous without ever turning against humanity.

Microsoft blocks kids under 13 from Copilot as it pushes age checks for AI

Microsoft blocks kids under 13 from Copilot as it pushes age checks for AI

Microsoft says children under 13 cannot use Copilot as it pushes an AI future built around age verification, differentiated experiences, and stronger protections for younger users.

OpenAI chief scientist warns humanity is not ready for what comes next

OpenAI chief scientist warns humanity is not ready for what comes next

OpenAI chief scientist Jakub Pachocki has issued an ominous warning about rapidly advancing artificial intelligence, saying nobody is prepared for what could happen as AI becomes smarter, harder to monitor, and increasingly capable of helping improve itself.

Mozilla says your open source AI might not be as open as you think

Mozilla says your open source AI might not be as open as you think

Mozilla is highlighting a new framework for understanding open source AI, arguing that model weights, training data, code, documentation, and other components can have very different levels of openness.

OpenAI wants to monitor AI misuse without reading your prompts

OpenAI wants to monitor AI misuse without reading your prompts

OpenAI wants to detect dangerous patterns across multiple AI interactions without letting its employees read customer prompts. Here is how Private Safety Processing and Zero Data Retention are supposed to work.

OpenAI says Astra could be dangerously good at cyberattacks

OpenAI says Astra could be dangerously good at cyberattacks

OpenAI says its upcoming Astra model may have reached a critical cybersecurity capability level. I am glad the company is focused on safety, but the repeated warnings about its own creations are getting tiring.