Building with AI
What is RLHF (Reinforcement Learning from Human Feedback)?
Training method that uses human preferences to make models more helpful and aligned.
RLHF fine-tunes a model using human rankings of its outputs, teaching it to prefer responses people find helpful and safe.
It's a major reason modern chat assistants feel useful and well-behaved.
Related terms
Find building with ai tools
Browse hand-reviewed AI tools — compare on pricing, features and real ratings.