Agents trust tools too much; measuring reliance on unreliable tools in tool-using agents
Read the original at arxiv.org→arXiv:2609.05587v1 Announce Type: new Abstract: Existing evaluations of tool-using agents primarily measure whether an agent can successfully complete diverse tasks with tools. These evaluations generally assume...
Original headline: "Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools"
Coverage timeline
- Sep 9, 04:00 UTC arXiv cs.AI lead source Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools