Go to example tasks
Transcribe an audio file to text with timestamps
Transcribes a public audio or video link (mp3, m4a, wav, ogg, mp4, webm) with an open-source Whisper model. Returns the full text, timed segments and ready subtitle files (SRT, VTT). The language is detected automatically. You must have the right to process the file.
Audio & Podcast Transcriber – MP3, WAV, RSS to Text, SRTtinlark/audio-podcast-transcriber
Episode
File
Language
Seconds
+8 fieldsTextNumberBooleanListObject
Input
Audio or video file links:https://upload.wikimedia.org/wikipedia/commons/3/30/LibriVox_-_Everrett_Copy_of_the_Gettysburg_Address_-_Michael_Scherer.ogg
Model:base
Spoken language:auto
Output formats:text+2
Output fields
Episode
File
Language
Seconds
Words
Model
Billed min
Published
Warnings
Status
Error code
Error
Sign up on Apify01
Create your Apify account to access the Audio & Podcast Transcriber – MP3, WAV, RSS to Text, SRT.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
