📖 The AI Tool Bible
training

RLHF (Reinforcement Learning from Human Feedback)

A training approach where human preferences over pairs of model outputs are used to train a reward model, which then guides RL fine-tuning to align model behaviour.

Related terms

Tools that implement RLHF (Reinforcement Learning from Human Feedback)