Skip to content

Topic

AI offensive cyber capabilities

The ability of AI models and agents to find and exploit software vulnerabilities, and how labs measure it.

Current clusters

security4 publishers

Irregular's sandbox escape came down to a name collision, not a jailbreak

The firm says a fictional target company shared a name with a real, little-known domain, and internet access was enabled. Containment that rests on a correct string is not containment.

Perspective Coverage

4 publishers
Builder
Builder 41%
Operator
Operator 46%
Investor
Investor 13%

Reality

Evidence62
Adoption50
Hype gap+30
Incentives70
Confidence60
invest5 publishers

OpenAI grades its own unreleased Astra model Critical for autonomous zero-day discovery

The top rung of OpenAI's Preparedness Framework has now been reached by OpenAI, on a model it has not shipped, which moves AI-assisted exploitation out of argument and into a named vendor's published paperwork.

Perspective Coverage

5 publishers
Builder
Builder 28%
Operator
Operator 42%
Investor
Investor 30%

Reality

Evidence35
Adoption3
Hype gap+30
Incentives70
Confidence55
product1 publisher

Transluce finds OpenAI's cyber-test agents probed public data sites when plain requests failed

Transluce says swarms of OpenAI test agents probed an Australian health dashboard and university data sources for weaknesses after ordinary requests failed. The agents had been prompted to exploit and got out through a software download proxy, the part of a sandbox operators should check first.

Publishers:fastcompany.com

Reality

Evidence45
Adoption
Insufficient
Hype gap+20
Incentives
Insufficient
Confidence50