From a multilingual streaming ASR backbone to Kenyan-language systems: data-centric adaptation of Nemotron 3.5 for Kikuyu, Dholuo, and Kalenjin
Read the original at arxiv.org→arXiv:2607.18912v1 Announce Type: new Abstract: Automatic speech recognition (ASR) for African languages is constrained by orthographic inconsistency, annotation artifacts, missing audio, speaker and domain...
Original headline: "From a Multilingual Streaming ASR Backbone to Kenyan-Language Systems: Data-Centric Adaptation of Nemotron 3.5 for Kikuyu, Dholuo, and Kalenjin"