security1 publisher
Judge model choice swings AI-Infra-Guard's false positive rate fifteenfold
Tencent's Zhuque Lab has put its AI asset scanner on GitHub for nothing. The skill auditor asks a language model whether code looks malicious, and the lab's own benchmark puts the false positive rate anywhere between 1.2 and 18.67 percent.
Publishers:helpnetsecurity.com
Reality
- Evidence58
- Adoption42
- Hype gap−12
- Incentives52
- Confidence55