15 posts

2. Security

Latest posts
Fable 5
Fable 5

Anthropic details Fable 5's cyber safeguards and a draft jailbreak severity framework, sorting cybersecurity uses into risk tiers and scoring techniques on four axes.

by AI-360
Fable 5 Returns
Fable 5 Returns

US export controls on Fable 5 and Mythos 5 have lifted. Anthropic details the Amazon-reported bypass, its fix, and a new Glasswing-backed framework for scoring AI jailbreak severity

by AI-360
OpenAI Launches Daybreak to Tackle Cybersecurity's "Patching Bottleneck"
OpenAI Launches Daybreak to Tackle Cybersecurity's "Patching Bottleneck"

OpenAI's Daybreak shifts cyber AI focus from finding flaws to fixing them, pairing an upgraded GPT-5.5-Cyber and Codex Security with a new vendor partner programme and an open-source patching drive.

by AI-360
Mindgard Finds ChatGPT Safeguards Easily Bypassed to Generate Graphic Imagery
Mindgard Finds ChatGPT Safeguards Easily Bypassed to Generate Graphic Imagery

Mindgard researcher Jim Nightingale says he was left "shaken, and in tears" after finding ChatGPT could be tricked into generating graphic violent and sexual images with minimal prompting.

by AI-360
OpenAI details method for predicting model misbehaviour before launch
OpenAI details method for predicting model misbehaviour before launch

The paper landed four days after Anthropic's Fable 5 was withdrawn

by AI-360
OpenAI Publishes Evaluation Playbook as Frontier Model Testing Comes Under Scrutiny
OpenAI Publishes Evaluation Playbook as Frontier Model Testing Comes Under Scrutiny

OpenAI Publishes Evaluation Playbook as Frontier Model Testing Comes Under Scrutiny

by AI-360
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Your link has expired. Please request a new one.
Great! You've successfully signed up.
Great! You've successfully signed up.
Welcome back! You've successfully signed in.
Success! You now have access to additional content.