stochastic-policy

Tag

Cards List
#stochastic-policy

QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning

arXiv cs.LG · 2026-05-19 Cached

Introduces QuantFPFlow, a reinforcement learning framework that uses quantum amplitude estimation to achieve a quadratic speedup in estimating the Fokker-Planck partition function for continuous control, improving exploration and avoiding local optima.

0 favorites 0 likes
#stochastic-policy

How Maximum Entropy makes Reinforcement Learning Robust

ML at Berkeley · 2021-07-26 Cached

This article explains how incorporating Shannon entropy into reinforcement learning objectives creates more robust agents capable of handling unexpected or adversarial changes in rewards and dynamics.

0 favorites 0 likes
← Back to home

Submit Feedback