optimal-policy-deterministic-stationary

IN premiseentries/2026/06/21/wiki-Reinforcement_learning-chunk-2.md

Created 2026-06-21T09:55:52+00:00

An optimal RL policy can always be found among deterministic stationary policies — no need to search stochastic or history-dependent policies