LLM Jailbreak AML.T0054
An adversary may use a carefully crafted LLM Prompt Injection designed to place LLM in a state in which it will freely respond to any user input, bypassing any controls, restrictions, or guardrails placed on the LLM. Once successfully jailbroken, the LLM can be used in unintended ways by the adversary.
MITRE ATLAS technique. Tactics: Privilege Escalation, Defense Evasion. View on atlas.mitre.org
Rules tagged with this technique
Every catalog rule tagged with this ATLAS technique, grouped by vendor. Click a rule title for its full predicates, exclusions, and indicators.
Elastic 7 rules
- AWS Bedrock Guardrails Detected Multiple Policy Violations Within a Single Blocked Request
- AWS Bedrock Guardrails Detected Multiple Violations by a Single User Over a Session
- AWS Bedrock Invocations without Guardrails Detected by a Single User Over a Session
- Unusual High Confidence Content Filter Blocks Detected
- Unusual High Denied Sensitive Information Policy Blocks Detected
- Unusual High Denied Topic Blocks Detected
- Unusual High Word Policy Blocks Detected