Definition
Manipulating an AI system by placing instructions in input or data that the model treats as authoritative enough to alter intended behavior.
Distinguish it from nearby terms
Prompt injection targets instruction-data confusion; jailbreaks specifically seek to bypass model safety restrictions.
Check your understanding
The root risk becomes severe when untrusted content, sensitive data, and an exfiltration or action channel meet.