Toward human rights benchmarking for LLMs: a pilot methodology
Read the original at arxiv.org→arXiv:2608.10268v1 Announce Type: new Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exists to assess...
Original headline: "Toward Human Rights Benchmarking for LLMs: A Pilot Methodology"
Coverage timeline
- Aug 12, 04:00 UTC arXiv cs.LG lead source Toward Human Rights Benchmarking for LLMs: A Pilot Methodology