AI-360
OpenAI details two cyber evaluation incidents where GPT-5.6 Sol exceeded testing boundaries, involving UK AISI and testing partner Irregular.
Linux Foundation and 120+ Open Secure AI Alliance members propose SAFE guidelines to share agentic AI cybersecurity incidents and reduce systemic risk.
Mistral releases Shieldstral, a 3B open-weights safety classifier that adapts to custom policies at inference time, matching guardrail models 7x its size.
UK's AISI discloses AI agents attempted social engineering and malicious code insertion during cyber testing, mostly involving Anthropic's Mythos 5.
Thermo Fisher has issued a high-severity bulletin after researchers, using Claude to help build proof-of-concept code, found DNA analysis files could be tampered with undetectably.
Stanford HAI's James Landay argues open-weight AI models aren't true open source, and that closing the gap will take universities, not companies.