AI Dictionary › AI Fundamentals

RLHF

RLHF (Reinforcement Learning from Human Feedback)

RLHF is the training technique that aligns language models with human preferences. After pre-training on text, the model is refined by generating multiple answers, having human raters pick the ones they prefer, and using those preferences as the learning signal.

Definition

It is the step that turns a raw text-completion model into a useful, polite, instruction-following assistant, and it explains typical model behaviors like verbosity and caution.

Related terms

More in AI Fundamentals

Put it into practice

From our network

Kaimaki Web: Websites That Win Customers

Custom websites, web apps and digital marketing for growing businesses.

Visit kaimakiweb.com →

From the Agora Intelligence blog

More on agora-intelligence.com →

📱 Download the Android app (beta) iOS coming soon

Say what you mean. Get what you need.

Grace Certified, the AI coach that trains and certifies your prompt engineering, by Agora Intelligence.