Back to feed
arXiv cs.CL·

Montreal Forced Aligner and the state of speech-to-text alignment in 2026

Signal
82
Hype
15
In three linesMontreal Forced Aligner 3.0, the reference tool since 2016 for forced speech-to-text alignment, achieves state-of-the-art performance on English, Japanese, and Korean with boundary errors <15ms. New capabilities: model adaptation, cross-language phone remapping, expanded language/dialect coverage, harmonized IPA dictionaries.
Read source
Your take?
VoiceBenchmarksOpen sourceTools

Summary generated by Claude — human-verified