Auditors fabricate: batch-size degradation and confident hallucination in LLM detection of planted document contamination
Read the original at arxiv.org→arXiv:2609.09696v1 Announce Type: new Abstract: Large language models are increasingly proposed as automated auditors of document quality, yet their reliability as detectors of planted errors is poorly...
Original headline: "When Auditors Fabricate: Batch-Size Degradation and Confident Hallucination in LLM Detection of Planted Document Contamination"
Coverage timeline
- Sep 10, 04:00 UTC arXiv cs.CL lead source When Auditors Fabricate: Batch-Size Degradation and Confident Hallucination in LLM Detection of Planted Document Contamination