YouTube Subtitles Scraper - Captions, Timestamps, Language avatar

YouTube Subtitles Scraper - Captions, Timestamps, Language

Pricing

$5.00 / 1,000 transcripts

Go to Apify Store
YouTube Subtitles Scraper - Captions, Timestamps, Language

YouTube Subtitles Scraper - Captions, Timestamps, Language

YouTube subtitles downloader: every caption line with start time and duration, the full text, and ready-made SRT and VTT files. Any language YouTube offers; human captions preferred over auto-generated. $5 per 1,000 videos, failed videos never charged. No login or API key.

Pricing

$5.00 / 1,000 transcripts

Rating

0.0

(0)

Developer

Khandji Omar

Khandji Omar

Maintained by Community

Actor stats

0

Bookmarked

1

Total users

0

Monthly active users

17 hours ago

Last modified

Share

YouTube Subtitles Scraper — Download Captions as Text with Timestamps

Download YouTube subtitles and closed captions for any video: every caption line with its start time and duration, plus the full text. Human-written subtitles are preferred; auto-generated ones are used only when nothing else exists.

What does this YouTube subtitles scraper do?

Give it YouTube video links or IDs. For every video it returns:

  • the full transcript as one clean string, ready for an LLM, a search index or a spreadsheet;
  • timestamped segments (start, duration, text) for subtitles, players or chunked RAG;
  • the language, whether captions are human-written or auto-generated, and every other caption language available;
  • the video title, channel, duration, word count and segment count.

Watch links, youtu.be short links, Shorts, embeds and bare 11-character IDs all work, mixed in one list.

Why use it

  • Timed subtitles — start and duration for every line, ready to convert to SRT or VTT.
  • Any caption language YouTube offers — and the list of all available languages.
  • Human vs auto-generated is flagged on every video.
  • No login, no API key; failed videos are never charged.

How to get YouTube transcripts in 3 steps

  1. Click Try for free and paste one or more YouTube links in YouTube videos.
  2. Optional: set a Preferred language (en, fr, es, de, …). Leave it empty for English, then the video's default.
  3. Click Start. Download the results as JSON, CSV, Excel, HTML or XML, or read them from the API.

Input

FieldTypeWhat it does
videosarrayYouTube watch URLs, youtu.be links, Shorts, embeds or 11-character video IDs. Up to 1,000 per run.
languagestringPreferred two-letter caption language. Human-written captions are always preferred over auto-generated ones.
includeSegmentsbooleantrue (default) adds timestamped segments; false returns plain text only.
translateTostringExperimental. Asks YouTube for a machine translation (fr, es…). YouTube often refuses automated translation requests; you then get the original language plus a translation_note, never a failed video.
outputFormatsarrayAny of srt, vtt, chunks: ready-to-use subtitle files and ~250-word RAG chunks with start/end seconds.
proxyConfigurationobjectUses Apify Proxy by default. YouTube blocks datacenter IPs, so leave it on.
maxTotalChargeUsdnumberHard spending cap for the run.
{
"videos": [
"https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"jNQXAC9IVRw"
],
"language": "en",
"includeSegments": true
}

Output

One dataset item per video:

[
{
"video_id": "dQw4w9WgXcQ",
"video_url": "https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"title": "Rick Astley - Never Gonna Give You Up (Official Video)",
"channel": "Rick Astley",
"language": "en",
"is_generated": false,
"available_languages": [
"en",
"de",
"es",
"fr",
"ja",
"pt"
],
"duration_seconds": 211.3,
"segment_count": 61,
"word_count": 374,
"text": "We're no strangers to love. You know the rules and so do I ...",
"segments": [
{
"text": "We're no strangers to love",
"start": 18.64,
"duration": 3.24
},
{
"text": "You know the rules and so do I",
"start": 22.64,
"duration": 4.32
}
],
"source": "youtube"
}
]
FieldDescription
video_id, video_urlThe video, normalised to a watch URL
title, channelVideo title and channel name
language, is_generatedCaption language and whether YouTube auto-generated it
available_languagesEvery caption language the video offers
textThe full transcript as one string
segments[{text, start, duration}], in seconds
duration_seconds, segment_count, word_countSize of the transcript

A run also writes RUN_SUMMARY.json to its key-value store: videos requested, delivered, charged, and one line per failed video with the reason.

How much does it cost?

$0.005 per transcript — you pay only for videos that actually come back with a transcript.

VideosCost
100$0.50
1,000$5.00
10,000$50.00
  • Failed videos are never charged: captions disabled, private, deleted or invalid links cost $0.
  • You set the ceiling: maxTotalChargeUsd (default $1) stops the run before it spends more.
  • Apify's free plan includes monthly platform credit, enough for hundreds of transcripts.
  • One transcript in every run of 2+ videos is free, so you can check the output on your own videos.

For comparison, the most-used transcript Actors on the Store charge $0.005–$0.01 per video.

Use cases

  • Translate or re-time subtitles for your own videos.
  • Language learning — bilingual reading from real videos.
  • Accessibility — captions for embeds and players.
  • Research — search what was said and when.

Use it from code (API)

Every Apify Actor is an API. Start a run and read the results with one call.

Python

from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("om_kh/youtube-transcript-scraper").call(run_input={"videos": ["https://www.youtube.com/watch?v=dQw4w9WgXcQ", "jNQXAC9IVRw"], "language": "en", "includeSegments": True})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item)

JavaScript

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('om_kh/youtube-transcript-scraper').call({
"videos": [
"https://www.youtube.com/watch?v=dQw4w9WgXcQ",
"jNQXAC9IVRw"
],
"language": "en",
"includeSegments": true
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);

One HTTP call (cURL) — runs the Actor and returns the dataset in the same response:

curl -X POST "https://api.apify.com/v2/acts/om_kh~youtube-transcript-scraper/run-sync-get-dataset-items?token=<YOUR_APIFY_TOKEN>" \
-H "Content-Type: application/json" \
-d '{"videos": ["https://www.youtube.com/watch?v=dQw4w9WgXcQ", "jNQXAC9IVRw"], "language": "en", "includeSegments": true}'

Integrations: n8n, Make, Zapier, Google Sheets

  • n8n — the official Apify node, operation Run Actor and get dataset, Actor om_kh/youtube-transcript-scraper.
  • Make — Apify › Run an Actor, then Apify › Get Dataset Items.
  • Zapier — Apify › Run Actor; map the dataset items into your next step.
  • Google Sheets, webhooks, Slack — add an integration from the Integrations tab of this Actor.
  • Schedules — run it every hour or day from Schedules in Apify Console.

Use it from an AI agent (MCP and x402)

Connect Claude Desktop, Cursor, VS Code or any MCP client to Apify's MCP server at https://mcp.apify.com and add this Actor, om_kh/youtube-transcript-scraper, as a tool — the agent passes the same JSON as the run input and gets the dataset back.

AI agents that pay with x402 can run it too, without an Apify account: the Actor charges only per result delivered, so an agent pays exactly for what it gets.

Tips and limits

  • No captions, no transcript. This Actor reads YouTube's own captions. A video with captions turned off returns a NO_TRANSCRIPT error for that video, is not charged, and the rest of the run continues.
  • Pick the language you want. If your language is missing, check available_languages in the output and run again with one of them.
  • Big lists: up to 1,000 videos per run; for a whole channel or playlist use YouTube Channel Transcripts Scraper - Every Video to Text.
  • Speed: videos are fetched 8 at a time; a typical video takes 1–3 seconds.

FAQ

Can I get SRT or VTT subtitle files? Yes. Set outputFormats to ["srt"], ["vtt"] or both: each result then carries a ready-to-save srt / vtt string built from the timestamped captions. No extra charge.

Is the output ready for RAG and LLMs? Yes. outputFormats: ["chunks"] adds chunks: consecutive captions merged into ~250-word pieces, each with start and end seconds, so you can embed them and cite the exact moment in the video.

Can I translate a transcript? There is an experimental translateTo option that asks YouTube for its own machine translation. YouTube often refuses automated translation requests; when it does you still get the original transcript and a translation_note (and pay only the normal price). For reliable translation, send the text or chunks to your own translation model.

Is it legal to scrape YouTube transcripts? This Actor only reads captions that YouTube shows publicly to any visitor. It does not log in and collects no personal data beyond the public channel name. You are responsible for how you use the text; check YouTube's terms and your local law, and consult a lawyer if unsure.

Does it need a YouTube API key or cookies? No. There is no key, no login and nothing that expires.

Can it transcribe videos that have no captions? No — it returns YouTube's existing captions (human or auto-generated), it does not run speech-to-text. Almost all spoken-word videos have auto-generated captions.

Which languages are supported? Every language YouTube offers captions in. Set language, or leave it empty for English first.

Can I get timestamps? Yes, includeSegments (on by default) returns every caption line with start and duration in seconds.

Am I charged for videos that fail? Never. Only transcripts that are actually delivered are charged.

Can I export to CSV or Excel? Yes, from the Storage tab of any run, or with ?format=csv on the dataset API.

Can my AI agent call it? Yes, via MCP (see above) or the plain HTTP API.

Support

Found a bug or need a feature? Open an issue in the Issues tab of this Actor — every report gets an answer, usually within a day.