Audio · Free · collected from HuggingFace
Visit Mms 300M 1130 Forced Aligner ↗
Mms 300M 1130 Forced Aligner is a high‑performance forced alignment model that aligns speech audio with transcript text. It uses a 300‑million parameter encoder trained on a 1130‑hour multilingual dataset, delivering accurate word‑level timing across many languages.
ai huggingface automatic-speech-recognition
Listed as Free. Pricing changes often — confirm on the official site.
Listings are collected automatically from public sources and refreshed daily. We do not take payment for placement.