Subtitles from audio, without a video file
Speech recognition works on the audio track, so a video is never actually required — the picture contributes nothing to the transcript. Uploading audio directly skips a step and produces a smaller file to transfer.
This matters most for podcasts, recorded interviews, lectures and voice notes, where a video version either does not exist or is added later. The subtitle file generated from audio will drop straight onto a video of the same recording, because the timings refer to the audio either way.




