security1 distinct publisher
OpenAI gates a 100% ExploitBench model behind refusals it plans to loosen in weeks
Astra scored a perfect 100% on OpenAI's own exploit-development benchmark, against 78.5% for GPT-5.6 Sol, and the shipped model's refusal to write proof-of-concept code is a policy the company has already said it will relax.
Publishers:thehackernews.com
Reality
- Evidence27
- Adoption21
- Hype gap+46
- Incentives84