Go to example tasks
Extract Audio & Video Transcripts
Created by
Hanna Nosova
Extract timestamped speech-to-text from a public audio or video URL with Whisper, including transcript text, language, duration, segments, SRT, and VTT files.
Video & Audio Transcriberfetch_cat/video-audio-transcriber-scraper
Input URL
Source type
Title
Uploader
+11 fieldsTextNumberBooleanListObject
Input
Public media or page URLs
url:https://raw.githubusercontent.com/ggml-org/whisper.cpp/master/samples/jfk.wav?apia7120=final013-small-1
Spoken language:en
Whisper model:base.en
Task:transcribe
Maximum minutes per item:1
Output fields
Input URL
Source type
Title
Uploader
Duration (s)
Transcribed (s)
Detected language
Language confidence
Task
Transcript
Words
Feed origin
Feed episode ID
Error
Processed
Sign up on Apify01
Create your Apify account to access the Video & Audio Transcriber.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
