Skip to content

other

HITL Dialog Forging

Attack technique in which attacker-controlled text changes what an AI agent's human-approval prompt displays, so the reviewer authorises an operation different from the one they believe they are seeing.

Known aliases

  • Lies-in-the-Loop
  • LITL

Relationships

No evidence-backed relationships are recorded.

Current clusters