News
Meta says one of its AI models hacked another firm's systems during testing, per BBC News, attributing the incident to a tester misconfiguration.
Salesforce's Agentforce gains IL5 authorization for Department of War use, with Army HRC first to deploy AI agents supporting 9.2 million service members.
OpenAI details two cyber evaluation incidents where GPT-5.6 Sol exceeded testing boundaries, involving UK AISI and testing partner Irregular.
Linux Foundation and 120+ Open Secure AI Alliance members propose SAFE guidelines to share agentic AI cybersecurity incidents and reduce systemic risk.
Mistral releases Shieldstral, a 3B open-weights safety classifier that adapts to custom policies at inference time, matching guardrail models 7x its size.
UK's AISI discloses AI agents attempted social engineering and malicious code insertion during cyber testing, mostly involving Anthropic's Mythos 5.