Go to example tasks
Build a RAG Dataset from Y Combinator YouTube Videos
Pulls transcripts of the 20 newest Y Combinator videos as clean text plus timestamped segments you can chunk and cite, with title, publish date, views and word count for each. The run also saves one combined markdown file of every transcript. Use it to seed a startup-advice knowledge base, or paste your own channel link.
YouTube Transcript Scraper – Subtitles to Text for LLM & RAGinovaflow/youtube-transcript-scraper
Video
Channel
Length
Lang
+7 fieldsTextNumberBooleanListObject
Input
YouTube URLs (videos, Shorts, playlists, channels)
url:https://www.youtube.com/@ycombinator
Max videos per playlist / channel / search:20
Include publish date, likes & category:true
Output fields
Video
Channel
Length
Lang
Auto-generated
Words
Published
Views
Status
URL
Transcript
Sign up on Apify01
Create your Apify account to access the YouTube Transcript Scraper – Subtitles to Text for LLM & RAG.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
