β All models Β· MMS forced alignment
use case Β· practical
Subtitles need timestamps; forced alignment makes them. Drop in any audio you have a transcript for, align it on-device, and export a ready-to-use SRT block β the exact format video players and editors consume. Audio never leaves your machine.
CC-BY-NC-4.0 β non-commercial weights. For personal/educational use; see the main page for full license + attribution.
Bundled JFK sample: inaugural address excerpt (20 January 1961), spoken by John F. Kennedy β public domain, U.S. government work. JFK audio source and attribution record.
SRT output (one caption per word β group them in your editor)
The SRT uses word-level captions so you can see the raw timings; most editors can merge them into sentence cues. Try it with your own recording and its transcript β the aligner only needs the words to match the audio.
β Back to MMS forced alignment Β· Next: Wild β