Memory Merge DQN: sensitivity weighted target updates for stable value learning
Read the original at arxiv.org→arXiv:2607.19397v1 Announce Type: new Abstract: Deep Q-networks use target networks to stabilise bootstrapped value learning, but the standard hard copy update also introduces a tradeoff. Holding the target network...
Original headline: "Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning"
Coverage timeline
- Jul 23, 04:00 UTC arXiv cs.LG lead source Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning