Benchmarking the Benchmarks: evaluating automated safety benchmarks for small language models
Read the original at arxiv.org→arXiv:2608.17183v1 Announce Type: new Abstract: Small Language Models (SLMs) are increasingly deployed in resource-constrained, privacy-sensitive settings, where safety and bias failures can cause security and...
Original headline: "Benchmarking the Benchmarks: Evaluating Automated Safety Benchmarks for Small Language Models"
Coverage timeline
- Aug 19, 04:00 UTC arXiv cs.AI lead source Benchmarking the Benchmarks: Evaluating Automated Safety Benchmarks for Small Language Models