Auditability is not one property: rule overlap, behavioural agreement, and composition in reinforcement learning
Read the original at arxiv.org→arXiv:2609.28581v1 Announce Type: new Abstract: Reinforcement-learning (RL) policies are often distributed as opaque neural checkpoints, while training logs show that a run occurred without explaining what the...
Original headline: "Auditability Is Not One Property: Rule Overlap, Behavioural Agreement, and Composition in Reinforcement Learning"
Coverage timeline
- Sep 26, 04:00 UTC arXiv cs.LG lead source Auditability Is Not One Property: Rule Overlap, Behavioural Agreement, and Composition in Reinforcement Learning