TL;DR Prompt injection is an attack where untrusted text, either typed by a user or hidden in content an AI application reads, is treated by a large language model (LLM) as instructions. This can make the model ignore its rules or take actions it shouldn’t. Prompt injection is ranked the...