Home Knowledge Base Reward Learning

Reward Learning is the general problem of learning a reward function from various forms of human feedback — including demonstrations, preferences, corrections, natural language instructions, and other signals, to define the objective for RL without manual reward engineering.

Forms of Reward Feedback

Why It Matters

Reward Learning is letting humans define the objective — automatically constructing reward functions from diverse forms of human feedback.

reward learningrlhf

Explore 500+ Semiconductor & AI Topics

From EUV lithography to CUDA optimization — search the full knowledge base or chat with our AI assistant.