Go to example tasks
Quick Start: Audio & Video to Text - Whisper Transcription
Ready-to-run starter configuration for Audio & Video to Text - Whisper Transcription API. Copy it, change the visible inputs, and get structured results quickly.
Audio & Video to Text - Whisper Transcription APIzenomastro/audio-video-to-text
Media URL
Status
Duration (s)
Billed minutes
+5 fieldsTextNumberBooleanListObject
Input
Audio or video file URLs(required):https://upload.wikimedia.org/wikipedia/commons/4/46/1941_Roosevelt_speech_pearlharbor_p1.ogg
Output formats:text+2
Maximum duration per file (minutes):5
Parallel transcriptions:1
Output fields
Media URL
Status
Duration (s)
Billed minutes
Language
Words
Transcript
SRT file
Error
Sign up on Apify01
Create your Apify account to access the Audio & Video to Text - Whisper Transcription API.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
