Evaluation of forced alignment of code-mixed speech: the case of Hindi-English

arXiv:2607.25581v1 Announce Type: new Abstract: Code-mixed speech poses unique challenges to forced alignment: expanded inventories, orthographic errors, and speaker variation. We evaluate forced alignment of Hindi-English code-mixed speech using the Montreal Forced Aligner. We address 2 problems: (1) free variation involving native vs non-native pairs and (2) phonemic boundary detection for mid-utterance English words. Bootstrapping strategies substantially outperform unmodified lexicons. Acous...

arXiv cs.CL ·Ayushi Pandey, Pamir Gogoi, Kevin Tang ·
compartilhar: