AI Dictionary › AI Fundamentals
RLHF (Reinforcement Learning from Human Feedback)
RLHF is the training technique that aligns language models with human preferences. After pre-training on text, the model is refined by generating multiple answers, having human raters pick the ones they prefer, and using those preferences as the learning signal.
It is the step that turns a raw text-completion model into a useful, polite, instruction-following assistant, and it explains typical model behaviors like verbosity and caution.
From our network
Kaimaki Web: Websites That Win Customers
Custom websites, web apps and digital marketing for growing businesses.
Visit kaimakiweb.com →From the Agora Intelligence blog
📱 Download the Android app (beta) iOS coming soon
Say what you mean. Get what you need.
Grace Certified, the AI coach that trains and certifies your prompt engineering, by Agora Intelligence.