Home › Glossary › Prompt Shields for user prompts

Prompt Shields for user prompts

See it in the glossary A–Z

The Prompt Shields check for direct attacks within the user's input, such as attempts to change rules, role-play personas, mocked-up conversations and encoding tricks. Documents and images are not scanned.

Also called user prompt attack, jailbreak, Jailbreak risk detection.

Read more: Microsoft Learn

In the Ultra Transcenders books

AI-103

Each book explains Prompt Shields for user prompts in context, with comparison tables and the common traps.

Terms in this definition

Related terms

See Prompt Shields for user prompts in the full glossary