YouTube Transcript Scraper - Full Text, Timestamps, Language avatar

YouTube Transcript Scraper - Full Text, Timestamps, Language

Pricing

Pay per usage

Go to Apify Store
YouTube Transcript Scraper - Full Text, Timestamps, Language

YouTube Transcript Scraper - Full Text, Timestamps, Language

YouTube transcript API for single videos and whole channels: full text, timestamped segments, views, likes, comments, upload date and channel subscribers. Optional SRT/VTT and RAG chunks. Up to 1,000 videos per run. Free until 12 Oct 2026, then $5 per 1,000; failed videos never charged.

Pricing

Pay per usage

Rating

0.0

(0)

Developer

Khandji Omar

Khandji Omar

Maintained by Community

Actor stats

1

Bookmarked

55

Total users

14

Monthly active users

4 days ago

Last modified

Share

YouTube Transcript Scraper — Full Text, Timestamps, Any Language

Get the transcript of any YouTube video as clean text plus timestamped segments. Paste links or video IDs, get one structured transcript per video — for summaries, research, subtitles, SEO content and AI pipelines. No API key, no login.

What does this YouTube transcript scraper do?

Give it YouTube video links or IDs. For every video it returns:

  • the full transcript as one clean string, ready for an LLM, a search index or a spreadsheet;
  • timestamped segments (start, duration, text) for subtitles, players or chunked RAG;
  • the language, whether captions are human-written or auto-generated, and every other caption language available;
  • the video title, channel, duration, word count and segment count.

Watch links, youtu.be short links, Shorts, embeds and bare 11-character IDs all work, mixed in one list.

Why use it

  • Reads YouTube's own captions — the exact text YouTube shows, not a speech-to-text guess.
  • Human captions first, auto-generated only when nothing better exists — and the output says which.
  • Clean output — one string for LLMs and timestamped segments, in the same row.
  • The video's numbers in the same row — views, likes, comment count, upload date, description, tags, channel and thumbnail, so you do not need a second scraper.
  • Up to 1,000 videos per run, fetched in parallel — paste video links, or whole channels (up to 500 videos each, with a date range).
  • Honest billing — failed videos are never charged; every failure is listed with its reason.
  • API, MCP and x402 — one HTTP call from code, a tool for AI agents, payable by agents with x402.

How to get YouTube transcripts in 3 steps

  1. Click Try for free and paste one or more YouTube links in YouTube videos.
  2. Optional: set a Preferred language (en, fr, es, de, …). Leave it empty for English, then the video's default.
  3. Click Start. Download the results as JSON, CSV, Excel, HTML or XML, or read them from the API.

Input

FieldTypeWhat it does
videosarrayYouTube watch URLs, youtu.be links, Shorts, embeds or 11-character video IDs. Up to 1,000 per run.
languagestringPreferred two-letter caption language. Human-written captions are always preferred over auto-generated ones.
includeSegmentsbooleantrue (default) adds timestamped segments; false returns plain text only.
channelsarrayWhole channels: links, @handles or channel IDs. Newest videos first. Use with or instead of videos.
maxVideosPerChannelinteger1 to 500 per channel. Default 10.
publishedAfter, publishedBeforestringKeep only channel videos uploaded in this date range (YYYY-MM-DD).
includeStatsbooleantrue (default) adds likes, comment count, upload date and the channel's subscribers, handle and avatar. Views, description, channel and thumbnail are always included.
translateTostringExperimental. Asks YouTube for a machine translation (fr, es…). YouTube often refuses automated translation requests; you then get the original language plus a translation_note, never a failed video.
outputFormatsarrayAny of srt, vtt, chunks: ready-to-use subtitle files and ~250-word RAG chunks with start/end seconds.
proxyConfigurationobjectUses Apify Proxy by default. YouTube blocks datacenter IPs, so leave it on.
maxTotalChargeUsdnumberHard spending cap for the run.
{
"videos": [
"https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"jNQXAC9IVRw"
],
"language": "en",
"includeSegments": true
}

Output

One dataset item per video:

[
{
"video_id": "dQw4w9WgXcQ",
"video_url": "https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"title": "Rick Astley - Never Gonna Give You Up (Official Video)",
"channel": "Rick Astley",
"language": "en",
"is_generated": false,
"available_languages": [
"en",
"de",
"es",
"fr",
"ja",
"pt"
],
"duration_seconds": 211.3,
"segment_count": 61,
"word_count": 374,
"view_count": 1823638919,
"like_count": 19473969,
"comment_count": 2400000,
"published_at": "2009-10-24",
"channel_id": "UCuAXFkgsw1L7xaCfnd5JJOw",
"channel_url": "https://www.youtube.com/channel/UCuAXFkgsw1L7xaCfnd5JJOw",
"channel_handle": "@RickAstleyYT",
"channel_subscribers": 4550000,
"description": "The official video for “Never Gonna Give You Up” by Rick Astley...",
"thumbnail_url": "https://i.ytimg.com/vi/dQw4w9WgXcQ/sddefault.jpg",
"video_length_seconds": 213,
"is_live": false,
"text": "We're no strangers to love. You know the rules and so do I ...",
"segments": [
{
"text": "We're no strangers to love",
"start": 18.64,
"duration": 3.24
},
{
"text": "You know the rules and so do I",
"start": 22.64,
"duration": 4.32
}
],
"source": "youtube"
}
]
FieldDescription
video_id, video_urlThe video, normalised to a watch URL
title, channelVideo title and channel name
language, is_generatedCaption language and whether YouTube auto-generated it
available_languagesEvery caption language the video offers
textThe full transcript as one string
segments[{text, start, duration}], in seconds
duration_seconds, segment_count, word_countSize of the transcript
view_count, like_count, comment_countVideo statistics. Likes are exact; YouTube rounds the comment count on very large videos
published_atUpload date (YYYY-MM-DD)
description, keywords, thumbnail_urlThe video's description, tags and best thumbnail
channel_id, channel_urlThe channel, ready to join with other data
channel_handle, channel_subscribers, channel_thumbnail_urlThe channel's @handle, subscriber count (as YouTube shows it, e.g. 8,680,000) and avatar
video_length_seconds, is_liveLength of the video itself, and whether it was a live stream

A run also writes RUN_SUMMARY.json to its key-value store: videos requested, delivered, charged, and one line per failed video with the reason.

How much does it cost?

$0.005 per transcript — you pay only for videos that actually come back with a transcript.

VideosCost
100$0.50
1,000$5.00
10,000$50.00
  • Failed videos are never charged: captions disabled, private, deleted or invalid links cost $0.
  • You set the ceiling: maxTotalChargeUsd (default $1) stops the run before it spends more.
  • Apify's free plan includes monthly platform credit, enough for hundreds of transcripts.

This Actor stays free until 12 October 2026; the price above applies from that date.

For comparison, the most-used transcript Actors on the Store charge $0.005–$0.01 per video.

Use cases

  • Summarise videos with ChatGPT or Claude — feed text straight into a prompt.
  • RAG and AI agents — chunk segments by timestamp and cite the exact moment.
  • SEO and content — turn videos into blog posts, show notes and quotes.
  • Market and competitor research — what reviewers and creators actually say.
  • Accessibility and subtitles — reuse the timed captions.

Use it from code (API)

Every Apify Actor is an API. Start a run and read the results with one call.

Python

from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("om_kh/youtube-transcript-api").call(run_input={"videos": ["https://www.youtube.com/watch?v=dQw4w9WgXcQ", "jNQXAC9IVRw"], "language": "en", "includeSegments": True})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item)

JavaScript

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('om_kh/youtube-transcript-api').call({
"videos": [
"https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"jNQXAC9IVRw"
],
"language": "en",
"includeSegments": true
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);

One HTTP call (cURL) — runs the Actor and returns the dataset in the same response:

curl -X POST "https://api.apify.com/v2/acts/om_kh~youtube-transcript-api/run-sync-get-dataset-items?token=<YOUR_APIFY_TOKEN>" \
-H "Content-Type: application/json" \
-d '{"videos": ["https://www.youtube.com/watch?v=dQw4w9WgXcQ", "jNQXAC9IVRw"], "language": "en", "includeSegments": true}'

Integrations: n8n, Make, Zapier, Google Sheets

  • n8n — the official Apify node, operation Run Actor and get dataset, Actor om_kh/youtube-transcript-api.
  • Make — Apify › Run an Actor, then Apify › Get Dataset Items.
  • Zapier — Apify › Run Actor; map the dataset items into your next step.
  • Google Sheets, webhooks, Slack — add an integration from the Integrations tab of this Actor.
  • Schedules — run it every hour or day from Schedules in Apify Console.

Use it from an AI agent (MCP and x402)

Connect Claude Desktop, Cursor, VS Code or any MCP client to Apify's MCP server at https://mcp.apify.com and add this Actor, om_kh/youtube-transcript-api, as a tool — the agent passes the same JSON as the run input and gets the dataset back.

AI agents that pay with x402 can run it too, without an Apify account: the Actor charges only per result delivered, so an agent pays exactly for what it gets.

Tips and limits

  • No captions, no transcript. This Actor reads YouTube's own captions. A video with captions turned off returns a NO_TRANSCRIPT error for that video, is not charged, and the rest of the run continues.
  • Pick the language you want. If your language is missing, check available_languages in the output and run again with one of them.
  • Big lists: up to 1,000 videos per run; for a whole channel or playlist use YouTube Channel Transcripts Scraper - Every Video to Text.
  • Speed: videos are fetched 8 at a time; a typical video takes 1–3 seconds.

FAQ

Can I get SRT or VTT subtitle files? Yes. Set outputFormats to ["srt"], ["vtt"] or both: each result then carries a ready-to-save srt / vtt string built from the timestamped captions. No extra charge.

Is the output ready for RAG and LLMs? Yes. outputFormats: ["chunks"] adds chunks: consecutive captions merged into ~250-word pieces, each with start and end seconds, so you can embed them and cite the exact moment in the video.

Can I translate a transcript? There is an experimental translateTo option that asks YouTube for its own machine translation. YouTube often refuses automated translation requests; when it does you still get the original transcript and a translation_note (and pay only the normal price). For reliable translation, send the text or chunks to your own translation model.

Is it legal to scrape YouTube transcripts? This Actor only reads captions that YouTube shows publicly to any visitor. It does not log in and collects no personal data beyond the public channel name. You are responsible for how you use the text; check YouTube's terms and your local law, and consult a lawyer if unsure.

Does it need a YouTube API key or cookies? No. There is no key, no login and nothing that expires.

Can it transcribe videos that have no captions? No — it returns YouTube's existing captions (human or auto-generated), it does not run speech-to-text. Almost all spoken-word videos have auto-generated captions.

Which languages are supported? Every language YouTube offers captions in. Set language, or leave it empty for English first.

Can I get timestamps? Yes, includeSegments (on by default) returns every caption line with start and duration in seconds.

Am I charged for videos that fail? Never. Only transcripts that are actually delivered are charged.

Can I export to CSV or Excel? Yes, from the Storage tab of any run, or with ?format=csv on the dataset API.

Can my AI agent call it? Yes, via MCP (see above) or the plain HTTP API.

Support

Found a bug or need a feature? Open an issue in the Issues tab of this Actor — every report gets an answer, usually within a day.