Memory-State Critic for asymmetric actor-critic with application to vision-based pursuit-evasion
Read the original at arxiv.org→arXiv:2610.03830v1 Announce Type: new Abstract: In partially observable Markov decision processes, the optimal policy generally depends on the history of observations and past actions. Asymmetric actor-critic...
Original headline: "Memory-State Critic for Asymmetric Actor-Critic with Application to Vision-Based Pursuit-Evasion"
Coverage timeline
- Oct 6, 04:00 UTC arXiv cs.LG lead source Memory-State Critic for Asymmetric Actor-Critic with Application to Vision-Based Pursuit-Evasion