Assessing how language disparity affects multilingual linguistic ability in large language models across 101 languages using the MultiBLiMP benchmark and four evaluation methods.
Read the original at arxiv.org→arXiv:2610.00540v1 Announce Type: new Abstract: Claims about the grammatical competence of multilingual language models vary sharply with how competence is measured, yet the interaction between evaluation paradigm,...
Original headline: "Assessing the Impact of Language Disparity on Multilingual Linguistic Ability in Large Language Models"
Coverage timeline
- Oct 2, 04:00 UTC arXiv cs.CL lead source Assessing the Impact of Language Disparity on Multilingual Linguistic Ability in Large Language Models