NEWFresh AI tools added every week. Explore what's trending across 60+ categories.See what's new →
Building with AI

What is RLHF (Reinforcement Learning from Human Feedback)?

Training method that uses human preferences to make models more helpful and aligned.

RLHF fine-tunes a model using human rankings of its outputs, teaching it to prefer responses people find helpful and safe.

It's a major reason modern chat assistants feel useful and well-behaved.

Related terms

← All AI terms · Browse AI tools

Find building with ai tools

Browse hand-reviewed AI tools — compare on pricing, features and real ratings.