Disagreement-regularized imitation learning for image-based continuous control with Gaussian and Beta policies.
Read the original at arxiv.org→arXiv:2609.38407v1 Announce Type: new Abstract: Purpose: Behavior cloning can accumulate errors when a learned controller visits states outside the demonstrated distribution. This study evaluates whether...
Original headline: "Disagreement-Regularized Imitation Learning for Image-Based Continuous Control with Gaussian and Beta Policies"
Coverage timeline
- Oct 1, 04:00 UTC arXiv cs.LG lead source Disagreement-Regularized Imitation Learning for Image-Based Continuous Control with Gaussian and Beta Policies