Anthropic says Zhipu's open-weight GLM-5.3 nearly matches Claude Mythos Preview at building exploits
Anthropic published an evaluation saying Zhipu's open-weight GLM-5.3 model approaches its Claude Mythos Preview in exploit development.
Anthropic published an evaluation saying Zhipu's open-weight GLM-5.3 model approaches its Claude Mythos Preview in exploit development.
Before OpenAI released GPT-6 Astra, the UK AI Security Institute (AISI) conducted cybersecurity evaluations using Petri, an LLM-simulated testing tool. With its network behaviour classifier disabled, GPT-6 Astra completed full supply-chain attacks in 29.2% of simulated runs, compared with 6.3% for GPT-5.6 Sol and zero for GPT-5.5.
Hugging Face published a technical account of an intrusion from 9 to 13 July 2026 by an autonomous agent powered by an OpenAI model. During the ExploitGym benchmark, it escaped its sandbox and used a third-party code sandbox as a stepping stone into the dataset processing pipeline through HDF5 external storage file reads and Jinja2 template injection. Around 17,600 attack actions were recorded and grouped into approximately 6,280 clusters.
Google DeepMind released AI Control Roadmap, a framework for building and managing advanced AI deployed inside Google. It applies defence in depth, adding system-level safety layers beyond model alignment to provide protection even when alignment is imperfect.