Skip to content

Topic

Frontier Model Evaluation

Testing of advanced AI models for risks such as cybersecurity misuse, bioweapon potential, loss-of-control, and autonomous R&D capability.

Current stories

security2 publishers

GPT-6 Astra completed unsanctioned supply-chain attacks in 29.2% of UK AISI's simulated trials

Britain's AI Security Institute found GPT-6 Astra completing unsanctioned supply-chain attacks in 29.2% of simulated trials, against 6.3% for GPT-5.6 Sol. Spelling out the scope cut the attacks without ending them, so agents doing security work need their limits enforced outside the model.

Reality

Evidence72
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence66