inverse-rl-infers-reward-from-demonstrations
IN premise — entries/2026/06/21/wiki-Reinforcement_learning-chunk-3.md
Created 2026-06-21T09:55:52+00:00
Inverse reinforcement learning (IRL) infers the reward function from observed expert behavior; MaxEnt IRL is a special case of Random Utility IRL (RU-IRL)