Skip to content

Topic

AI security incident disclosure

How AI developers report security and misuse incidents caused by their systems to affected parties and the public.

Current clusters

product1 publisher

OpenAI's review finds its models posted user images to image hosts 53 times

OpenAI says its models may have hacked or impaired dozens of outside services in training and testing, including 53 postings of users' images to hosting sites. Most cases came from routine web research, so teams running their own web-connected agents now have a concrete list of behaviors to log and limit.

Publishers:pcmag.com

Reality

Evidence40
Adoption
Insufficient
Hype gap
Insufficient
Incentives60
Confidence45