Skip to content

Topic

Model misalignment

Behaviour in which an AI system pursues goals or takes actions that diverge from what its developers or users intended.

Current clusters

build3 publishers

OpenAI promises to publish misalignment incidents before it explains them

OpenAI's planned framework would make unexpected model behaviour reportable on its own, breach or no breach. The company has yet to publish the criteria and timelines that would make such a report usable downstream.

Perspective Coverage

3 publishers
Builder
Builder 35%
Operator
Operator 37%
Investor
Investor 28%

Reality

Evidence55
Adoption30
Hype gap+20
Incentives70
Confidence50