Jagged judges: Epistemic stability under silence, pressure, and persistence
Read the original at arxiv.org→arXiv:2608.12645v1 Announce Type: new Abstract: LLM judges have become central infrastructure for model evaluations, online grading, and reward modeling. Judges are typically validated by accuracy on golden data,...
Original headline: "Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence"
Coverage timeline
- Aug 15, 04:00 UTC arXiv cs.AI lead source Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence