Fast models, slow evidence: a paired and self-audited evaluation of system-1 decision models for LLM agent harnesses
Read the original at arxiv.org→arXiv:2610.02267v1 Announce Type: new Abstract: Agent harnesses make many small, typed decisions per task: which model to call, which tool to use, whether retrieved text is relevant, whether an input carries an...
Original headline: "Fast Models, Slow Evidence: A Paired and Self-Audited Evaluation of System-1 Decision Models for LLM Agent Harnesses"
Coverage timeline
- Oct 5, 04:00 UTC arXiv cs.AI lead source Fast Models, Slow Evidence: A Paired and Self-Audited Evaluation of System-1 Decision Models for LLM Agent Harnesses