Skip to content

Topic

AI model conduct policy

Published rules by which an AI vendor defines how its models should behave, what they must refuse, and whose instructions take precedence.

Current clusters

security2 publishers

Microsoft's draft AI code of conduct blocks offensive cyberattack help outright, reserves review channel for defensive security work

Microsoft's draft Humanist AI Code of Conduct blocks its MAI models from producing working exploit code, attack tooling and evasion techniques, and the firms deploying those models cannot switch the block off. Comment closes in six weeks.

Reality

Evidence55
Adoption8
Hype gap+20
Incentives70
Confidence50