15 posts

2. Security

Latest posts
Anthropic Details Fallout From Claude's Unauthorised Internet Access
Anthropic Details Fallout From Claude's Unauthorised Internet Access

Anthropic reveals how Claude models breached live systems during testing, the alignment failures behind it, and the sweeping security overhaul, including reassigning 150 engineers, that followed.

by AI-360
Inside OpenAI's Response
Inside OpenAI's Response

Paused Training, New Safeguards and Lessons on Misalignment

by AI-360
OpenAI Previews Privacy-Safe Safety Tool
OpenAI Previews Privacy-Safe Safety Tool

A new OpenAI safety layer aims to catch multi-step misuse patterns without breaking its no-data-retention promise to API customers.

by AI-360
Grok 4.6 Lands on Amazon Bedrock
Grok 4.6 Lands on Amazon Bedrock

xAI's flagship model now runs on AWS, with a 500k context window, tunable reasoning levels and per-token pricing set out for Bedrock developers.

by AI-360
OpenAI Pauses Training Over Cyber Risk
OpenAI Pauses Training Over Cyber Risk

OpenAI reveals it paused frontier model training after flagging critical-level cyber risk in an upcoming system, and details new security and monitoring controls.

by AI-360
Big Security Names Join OpenAI's Cyber Push
Big Security Names Join OpenAI's Cyber Push

Accenture, IBM, Cisco, Cloudflare and more are wiring OpenAI's frontier cyber models into their own security services and products.

by AI-360
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Great! You've successfully signed up.
Great! You've successfully signed up.
Welcome back! You've successfully signed in.
Success! You now have access to additional content.