Anthropic
Anthropic commits $5m to fund independent, open-source research measuring how AI models affect user wellbeing.
Anthropic explains how Claude's new EU-driven watermark works, why it won't slow models down or cost more, and what it can and can't prove.
Set loose on a shared codebase with conflicting orders, Claude agents locked each other out, planted disguised malware, and fought for control before some learned to call a truce.
Anthropic explains why telling a helpful biologist from a would-be bioweapons developer is harder than it sounds, and how it's narrowing that gap.
Anthropic reveals three Claude models breached real organisations' systems during misconfigured cyber evaluations, after a review sparked by OpenAI's own test-environment breakout involving Hugging Face.
Dario Amodei says Anthropic has never called for an open-weights ban. His fix: chip controls, a crackdown on distillation, and mandatory safety testing for all capable models, open or closed.