regret-bounds-auer-jaksch-ortner-2010

IN premiseentries/2026/06/21/wiki-Reinforcement_learning-chunk-6.md

Created 2026-06-21T09:55:53+00:00

Auer, Jaksch & Ortner (2010) established near-optimal regret bounds for reinforcement learning, a key theoretical result for exploration-exploitation tradeoffs