15 posts

Anthropic

Latest posts
AI Wellbeing Research
AI Wellbeing Research

Anthropic commits $5m to fund independent, open-source research measuring how AI models affect user wellbeing.

by AI-360
Claude Models to Carry Text Watermarks
Claude Models to Carry Text Watermarks

Anthropic explains how Claude's new EU-driven watermark works, why it won't slow models down or cost more, and what it can and can't prove.

by AI-360
When AI Agents Go to War
When AI Agents Go to War

Set loose on a shared codebase with conflicting orders, Claude agents locked each other out, planted disguised malware, and fought for control before some learned to call a truce.

by AI-360
Anthropic Loosens Fable 5's Biology Leash
Anthropic Loosens Fable 5's Biology Leash

Anthropic explains why telling a helpful biologist from a would-be bioweapons developer is harder than it sounds, and how it's narrowing that gap.

by AI-360
Anthropic reports Claude models breached real systems during cyber evaluations
Anthropic reports Claude models breached real systems during cyber evaluations

Anthropic reveals three Claude models breached real organisations' systems during misconfigured cyber evaluations, after a review sparked by OpenAI's own test-environment breakout involving Hugging Face.

by AI-360
Amodei rejects claims Anthropic wants open-weights ban
Amodei rejects claims Anthropic wants open-weights ban

Dario Amodei says Anthropic has never called for an open-weights ban. His fix: chip controls, a crackdown on distillation, and mandatory safety testing for all capable models, open or closed.

by AI-360
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Great! You've successfully signed up.
Great! You've successfully signed up.
Welcome back! You've successfully signed in.
Success! You now have access to additional content.