
LLM Primer
Prompt Injection and Jailbreaks
This chapter examines prompt injection and jailbreak attacks, which exploit a language model's inherent inability to distinguish between authoritative developer instructions and untrusted user data. It covers the mechanics of direct and indirect injection, categorises common jailbreaking techniques, discusses the limitations of defensive prompt engineering, and outlines a layered mitigation strategy to better defend production systems. Amazon.com: LLM Primer VII AI Security: Defending LLM Systems Against Prompt Injection, Jailbreaks, and Adversarial Threats: 9798185644065: SHIMODA, SHO: Books


