PPO-HSC: An exploratory reinforcement learning framework based on wide-area policy coverage optimization
Read the original at arxiv.org→arXiv:2607.16206v1 Announce Type: new Abstract: This paper introduces PPO-HSC (Proximal Policy Optimization with High-order Sampling Coverage), an exploratory reinforcement learning framework designed to address the...
Original headline: "PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization"