Beyond the pale: assessing prevalence and contents of extremist speech in LLM training data
Read the original at arxiv.org→arXiv:2608.14813v1 Announce Type: new Abstract: Despite a strong interest on the part of the research community in the topic of trustworthy and safe AI, the composition of the text corpora that large language models...
Original headline: "Beyond the pale: Assessing prevalence and contents of extremist speech in LLM training data"
Coverage timeline
- Aug 18, 04:00 UTC arXiv cs.CL lead source Beyond the pale: Assessing prevalence and contents of extremist speech in LLM training data