NewNew: The enterprise guide to Agentic AI — 24 min read.

Read →
AI Glossary · Foundations

Reinforcement Learning from Human Feedback (RLHF)

A training technique that uses human preference judgments to align model behavior with desired norms.

Definition

What is Reinforcement Learning from Human Feedback (RLHF)?

Reinforcement Learning from Human Feedback (RLHF) is a training technique that uses human preference judgments to align model behavior with desired norms.

Category
Foundations
Glossary set
28 related terms
Audience
Enterprise AI leaders

Why does Reinforcement Learning from Human Feedback (RLHF) matter in enterprise AI?

Reinforcement Learning from Human Feedback (RLHF) matters in enterprise AI programs because it helps business and technology leaders align vocabulary, scope, ownership, and measurable outcomes.