papersSEP 10 04:00 UTC
Paper Establishes Optimality Guarantees for Nonsmooth H-Infinity Policy Search in Control
A new arXiv preprint examines policy search for continuous-time H-infinity control using full-order dynamic output feedback, a setting that is both nonconvex and nonsmooth. The authors develop algorithms with provable optimality guarantees, addressing the shortage of rigorous theoretical foundations for direct policy search in reinforcement learning and continuous control.