Reinforcement Learning from Human Feedback (RLHF)
RLHF trains an AI system using human judgments about which model outputs are more helpful, safe, or appropriate.
Explore AIstify's latest reporting, research, and expert analysis tagged with "human feedback", collected in one continuously updated archive.
RLHF trains an AI system using human judgments about which model outputs are more helpful, safe, or appropriate.
Human evaluation uses people to judge AI outputs on qualities that automated metrics cannot measure reliably.