Learning to Decide, Not to Reason: Parameter-Efficient Decision Operators via Low-Rank Activation Steering
Read the original at arxiv.org→arXiv:2610.06950v1 Announce Type: new Abstract: Injecting skills into a frozen language model currently costs a million parameters and a reinforcement-learning pipeline. We introduce \method{}, a System-1 decision...
Coverage timeline
- Oct 7, 04:00 UTC arXiv cs.LG lead source Learning to Decide, Not to Reason: Parameter-Efficient Decision Operators via Low-Rank Activation Steering