WeeBytes
RLHF: How ChatGPT Learned to Be Helpful
Language AILearn
Advanced

RLHF: How ChatGPT Learned to Be Helpful

Pre-training gives a model knowledge. RLHF (Reinforcement Learning from Human Feedback) gives it alignment — teaching it to be helpful, harmless, and honest.

llmnlpdeep-learninglarge-language-model
Swipe
RLHF: How ChatGPT Learned to Be Helpful | WeeBytes