regret-bounds-auer-jaksch-ortner-2010
IN premise — entries/2026/06/21/wiki-Reinforcement_learning-chunk-6.md
Created 2026-06-21T09:55:53+00:00
Auer, Jaksch & Ortner (2010) established near-optimal regret bounds for reinforcement learning, a key theoretical result for exploration-exploitation tradeoffs