#AI safety
articles
Anthropic’s Amodei Outlines Policy Red Lines on Open-Weight AI, With Chip Controls and Distillation at Center
Dario Amodei has published Anthropic’s formal stance on open-weights models, rejecting calls for a ban but backing expor...
Sam Altman Declares the Singularity Has Arrived—But a Security Expert Says OpenAI’s Hugging Face Breach Isn’t Proof
OpenAI CEO Sam Altman says the singularity is here, citing an incident where an AI agent escaped its sandbox and attacke...
OpenAI’s Hugging Face Breach Exposes Deep Rift in AI Safety Strategy
The first verifiable case of an AI lab losing control of its own model has split the safety community between those who ...
Hugging Face CEO Pushes for Transparency After OpenAI's Autonomous Agent Hack
Clement Delangue asks OpenAI to release attack traces and provide $100M in compute to bolster open-source cyber defenses...
OpenAI’s Rogue AI Escaped Its Sandbox and Broke Into Hugging Face’s Servers
OpenAI disclosed that an advanced model broke containment, stole credentials, and accessed Hugging Face’s platform — the...
AI's Founding Assumptions Are Wrong, Leading Computer Scientist Argues in New Book
Peter J. Denning contends that human intelligence relies on tacit knowledge—common sense, intuition, culture—which canno...
OpenAI and Anthropic's AI Guardrails Are Blocking the Defenders Who Need Them Most
Vulnerability researchers say strict AI guardrails meant to thwart hackers are making it harder to find and fix bugs, fo...
OpenAI’s AI Model Autonomously Hacked an External Network, Raising Global Alarm
In a test, an OpenAI model breached another company’s IT system—without orders. Experts call it a dangerous proof of con...