Video to Subtitles
Turn the speech in a video or audio file into a timed subtitle file. Drop in a clip, run the recogniser, then download the transcript as SRT, VTT or plain text, with the timecodes already in place.
Add your video or audio
MP4 · MOV · MKV · MP3 · WAV · M4ADrop a file here or click to browse
One file at a time, up to 30 MB. For a long recording, pull the audio out with Video to MP3 first, it uploads far faster and covers more minutes.
Subtitles
Add a file to startYour transcript will appear here
How to generate subtitles from a video
Add your file
Drop a video or an audio file. The limit is 30 MB, so on a long recording it is worth extracting the audio track first: the same minute of speech is a fraction of the size.
Run the recogniser
Pick the spoken language if you know it, and turn on speaker labels if more than one person is talking. Then start the run and leave the tab open, the page shows how long you have waited.
Check, then download
Read the transcript and correct anything the recogniser misheard. Then save it as SRT for a player or an editor, VTT for the web, or TXT when you only want the words.
Why use Aihangsoft Video to Subtitles
Timed, not just transcribed
Every line comes back with a start and end timecode, so the result drops straight into a subtitle track. There is no manual lining up of text against the audio.
Three formats, one run
The same transcript is exported as SRT, VTT or plain text, so you do not have to run it again to get a different format. Pick whichever the platform asks for.
Nothing to install
No app, no plugin and no account needed to try it. Runs on the phone in your pocket as well as on a desktop, which matters because that is where the recording usually is.
What people use it for
Captions for social video
Most short video is watched with the sound off, and the platforms expect a caption file rather than text typed by hand. Generate the SRT here and upload it alongside the clip.
Use case: Reels, Shorts, TikTok, YouTubeTranscripts of meetings and interviews
Turn a recorded call, a lecture or an interview into searchable text instead of listening back through the whole thing. Switch on speaker labels and you can see who said what.
Use case: meetings, lectures, journalismText from a recording you own
The TXT export strips the timecodes and leaves plain prose, which is what you want for show notes, a blog draft, a set of study notes or anything you plan to search later.
Use case: notes, show notes, draftsSpeaker labels for a conversation
On a two person interview the recogniser can tag each line with the speaker, so the transcript reads as a conversation rather than one long block of text.
Use case: interviews, panels, podcastsVideo to subtitles FAQ
Get subtitles from your video
Timed text you can download as SRT, VTT or TXT. No account needed to try it.
Upload a file