The reporting

1 articles

Topic: dialect inclusion

A studio microphone on a stand in a recording room

NEWS AI Policy & Access

NVIDIA fine-tuning cuts Saudi dialect speech errors from 55% to 30% word error rate in Najdi and Hijazi

NVIDIA's own benchmarks showed its multilingual speech model missed more than half the words in Najdi and Hijazi speech. A fine-tuning recipe using 133.7 hours of curated SADA 2022 audio cut that error rate nearly in half, with lessons for other underrepresented dialects.