Vector Symbolic Policy Gradient introduces a discrete-action actor that scores actions by hypervector similarity to the encoded state, with updates equivalent to advantage-weighted hypervector bundling and normalization under softmax policy gradient
Read the original at arxiv.org→arXiv:2608.18404v1 Announce Type: new Abstract: We answer this question with Vector-Symbolic Policy Gradient (VSPG), a discrete-action actor that represents each action by a unit-norm hypervector and scores it by...
Original headline: "Vector Symbolic Policy Gradient"
Coverage timeline
- Aug 20, 04:00 UTC arXiv cs.LG lead source Vector Symbolic Policy Gradient