Go to example tasks
Transcribe an interview with speaker labels (diarization)
Transcribe the famous Apollo 13 "Houston, we've had a problem" exchange and label who speaks when (Speaker 1, Speaker 2) with timestamps and a speaker count, using Deepgram nova-3. Use it for interviews, podcasts with guests, meetings and call recordings where you need to know who said what.
Audio & Video to Text - Speech to Text Transcription, SRTtidytools/audio-transcriber
Input
Input index
Audio
Media found on page
+14 fieldsTextNumberBooleanListObject
Input
Audio or video file URLs:https://upload.wikimedia.org/wikipedia/commons/6/6f/Apollo13-wehaveaproblem.ogg
Speaker labels (who said what):true
Vocabulary hints (optional):Houston, Apollo, Odyssey, Aquarius
Output fields
Input
Input index
Audio
Media found on page
Episode
Success
Source
Language
Duration seconds
Billed min
Word count
Speakers
Text
Summary
SRT
Charged
Error
Error type
Sign up on Apify01
Create your Apify account to access the Audio & Video to Text - Speech to Text Transcription, SRT.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
