Towards A Unified Policy Abstraction Theory and Representation Learning Approach in Markov Decision Processes

M Zhang, H Tang, J Hao, Y Zheng - arXiv preprint arXiv:2209.07696, 2022 - arxiv.org
Lying on the heart of intelligent decision-making systems, how policy is represented and
optimized is a fundamental problem. The root challenge in this problem is the large scale …