Proxy confidence: auditing black-box LLM agents with a surrogate's log-probabilities
Read the original at arxiv.org→arXiv:2610.03894v1 Announce Type: new Abstract: A deployed LLM agent emits tool calls, queries, and code that can be silently wrong -- by the time the error surfaces, the action has run. Frontier chat APIs hide the...
Original headline: "Proxy Confidence: Auditing Black-Box LLM Agents with a Surrogate's Log-Probabilities"
Coverage timeline
- Oct 6, 04:00 UTC arXiv cs.AI lead source Proxy Confidence: Auditing Black-Box LLM Agents with a Surrogate's Log-Probabilities