← Back to the AI Glossary 📖 Training & Fine-Tuning

Reinforcement Learning from Human Feedback (RLHF)

A training technique where human reviewers rate a model's responses, and the model is adjusted to produce more of what humans rated highly.

RLHF is a major reason modern chatbots feel helpful and conversational rather than just technically correct — it's specifically tuning for what humans actually prefer.
One of 60 free AI glossary terms
Plain-language definitions for the AI jargon you'll actually run into — no email needed, ever, for this section.
Browse the Full Glossary →