IntegrityBench benchmarks LLMs on misconduct classification, ethical action reasoning, and artifact-grounded decision making across 36 paired tasks under pressure; evaluates 18 frontier model variants
Read the original at arxiv.org→arXiv:2608.12345v1 Announce Type: new Abstract: Language models are increasingly deployed as co-scientists, yet their ability to uphold research integrity under institutional pressure remains unmeasured. We...
Original headline: "Diagnostic Foundation for Evaluating LLMs' Research Integrity as Co-Scientists"
Coverage timeline
- Aug 15, 04:00 UTC arXiv cs.AI lead source Diagnostic Foundation for Evaluating LLMs' Research Integrity as Co-Scientists