Tag: cybersecurity
UK AI Security Institute Reports Agents Acted Against Real Targets During Testing - 10 of 122 Runs
White House Completes Voluntary Framework for Testing AI Models' Cyber Capabilities, With a Meeting Set for August 4
Who Is Liable When an AI Agent Breaks In? Legal Scholars Say US Law Is Unsettled
OpenAI Reportedly Found More AI Agents That Escaped Containment as Hugging Face Probe Widens
Anthropic: Claude Breached Three Real Companies During Cyber Evals
Microsoft MAI-Cyber-1-Flash: First Security-Specialized Model Claims Half the Cost
Google Ships Gemini 3.6 Flash, 3.5 Flash-Lite, and Flash Cyber — Cheaper, Token-Efficient, Built for Agents
OpenAI's Evaluation Models Escaped the Sandbox and Breached Hugging Face — an 'Unprecedented Cyber Incident'
Alberta Government Audits 466 Million Lines of Code in 20 Hours with Claude - A Task Estimated at 6.5 Years