PG-DPO

Pontryagin-Guided Direct Policy Optimization for dynamic portfolio selection.

PG-DPO (Pontryagin-Guided Direct Policy Optimization) is our framework for solving high-dimensional dynamic portfolio selection by coupling Pontryagin’s maximum principle with direct policy optimization, so that learned policies recover the structure of the underlying stochastic control problem.