Towards A Unified Policy Abstraction Theory and Representation Learning Approach in Markov Decision Processes
Lying on the heart of intelligent decision-making systems, how policy is represented and
optimized is a fundamental problem. The root challenge in this problem is the large scale …
optimized is a fundamental problem. The root challenge in this problem is the large scale …