Reinforcement Learning from Human Feedback (RLHF)
RLHF trains an AI system using human judgments about which model outputs are more helpful, safe, or appropriate.
Explore AIstify's latest reporting, research, and expert analysis tagged with "language model training", collected in one continuously updated archive.
RLHF trains an AI system using human judgments about which model outputs are more helpful, safe, or appropriate.
Cross-entropy measures the difference between a predicted probability distribution and the correct target distribution.