Simulation-free reinforcement learning with Wasserstein-tilted flow maps
Read the original at arxiv.org→arXiv:2609.27033v1 Announce Type: new Abstract: Reward fine-tuning aims to update a pre-trained flow-based generative model to improve the downstream reward of its generated samples. Existing methods typically...
Original headline: "WTF?! Simulation-Free Reinforcement Learning with Wasserstein-Tilted Flow Maps"
Coverage timeline
- Sep 24, 04:00 UTC arXiv cs.LG lead source WTF?! Simulation-Free Reinforcement Learning with Wasserstein-Tilted Flow Maps