# YouTube Video Transcription - Video to Text at Scale (`vonsensey/youtube-video-transcription`) Actor

Get the full spoken text of any YouTube video, formatted like something a person would read: real paragraphs, grouped under the video's own chapters, with a timestamp on every line. Works on single videos or an entire channel, and you only pay for the transcripts you receive.

- **URL**: https://apify.com/vonsensey/youtube-video-transcription.md
- **Developed by:** [Blackcube](https://apify.com/vonsensey) (community)
- **Categories:** AI, Videos, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $5.00 / 1,000 transcripts

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## YouTube Video Transcription - Video to Text at Scale

<table style="border-collapse:collapse;width:100%;margin:0 0 4px">
<tr><td colspan="3" style="padding:9px 12px;background:#0B6E75;border:1px solid #0B6E75"><span style="color:#FFFFFF;font-weight:700;font-size:13px;letter-spacing:.3px">YouTube Transcript Suite</span><span style="color:#CFF0F2;font-size:12px"> &nbsp;&bull;&nbsp; 9 Actors, one codebase, one transcript billed once</span></td></tr>
<tr><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-transcript-scraper" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Transcript Scraper</a><br><span style="color:#6B7280;font-size:11px">RAG-Ready Text &amp; Subtitles</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-transcript-api" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Transcript API</a><br><span style="color:#6B7280;font-size:11px">Bulk Video Transcripts, Fast</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-subtitle-downloader" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Subtitle Downloader</a><br><span style="color:#6B7280;font-size:11px">SRT &amp; VTT Subtitle Files</span></td></tr>
<tr><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-shorts-transcript-scraper" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Shorts Transcript Scraper</a><br><span style="color:#6B7280;font-size:11px">Bulk Shorts to Text</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#ECFDF5;vertical-align:top;width:33%"><span style="color:#0B6E75;font-weight:700;font-size:13px">YouTube Video Transcription</span><br><span style="color:#0B6E75;font-size:11px;font-weight:600">&#10148; You are here</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-channel-transcript-scraper" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Channel Transcript Scraper</a><br><span style="color:#6B7280;font-size:11px">Whole Channel to Text</span></td></tr>
<tr><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-playlist-transcript-scraper" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Playlist Transcript Scraper</a><br><span style="color:#6B7280;font-size:11px">Whole Playlist to Text</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-closed-captions-extractor" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Closed Captions Extractor</a><br><span style="color:#6B7280;font-size:11px">Every Caption Track</span></td><td style="padding:9px 12px;border:1px solid #E5E7EB;background:#FFFFFF;vertical-align:top;width:33%"><a href="https://apify.com/vonsensey/youtube-transcript-rag-dataset" style="color:#111827;font-weight:700;font-size:13px;text-decoration:none">YouTube Transcript RAG Dataset Builder</a><br><span style="color:#6B7280;font-size:11px">Chunk-Ready Text</span></td></tr>
</table>

**More from this account:** [Website Contact & Email Suite](https://apify.com/vonsensey/website-contact-email-extractor) · [Career Site & ATS Jobs Suite](https://apify.com/vonsensey/career-page-job-postings-scraper-api) · [Google News Suite](https://apify.com/vonsensey/google-news-scraper-api) · [Keyword Research Suite](https://apify.com/vonsensey/google-keyword-ideas-scraper) · [Shopify Store Intelligence Suite](https://apify.com/vonsensey/shopify-store-leads-scraper) · [eBay Data Suite](https://apify.com/vonsensey/ebay-scraper-api) · [Amazon Reviews Suite](https://apify.com/vonsensey/amazon-reviews-scraper-api) · [Reddit](https://apify.com/vonsensey/reddit-scraper-posts-comments-api) · [Meta Ad Library](https://apify.com/vonsensey/facebook-ads-library-scraper-meta-ad-api) · [Vinted](https://apify.com/vonsensey/vinted-scraper-api) · [Trustpilot Review Intelligence Suite](https://apify.com/vonsensey/trustpilot-reviews-scraper) · [App Store & Google Play Reviews Suite](https://apify.com/vonsensey/app-store-google-play-reviews-scraper) · [Etsy Research Suite](https://apify.com/vonsensey/etsy-scraper) · [YouTube Comments Intelligence Suite](https://apify.com/vonsensey/youtube-comments-scraper) · [Telegram Channel Intelligence Suite](https://apify.com/vonsensey/telegram-channel-scraper-api) · [Business Reviews Suite](https://apify.com/vonsensey/business-reviews-aggregator-scraper-api) · [Google Trends Suite](https://apify.com/vonsensey/google-trends-scraper) · [Amazon Product Data Suite](https://apify.com/vonsensey/amazon-product-scraper-api) · [Nordic Marketplaces Suite](https://apify.com/vonsensey/dba-dk-scraper-api) · [Google Sheets Suite](https://apify.com/vonsensey/google-sheets-scraper-api) · [TikTok Suite](https://apify.com/vonsensey/tiktok-scraper-api) · [Snapchat Suite](https://apify.com/vonsensey/snapchat-scraper-api) · [LinkedIn Public Data Suite](https://apify.com/vonsensey/linkedin-scraper-api) · [Instagram Suite](https://apify.com/vonsensey/instagram-posts-scraper-api) · [Contact Validation Suite](https://apify.com/vonsensey/email-verifier-validator-api) · [Google Maps Suite](https://apify.com/vonsensey/google-maps-reviews-scraper-api) · [X / Twitter](https://apify.com/vonsensey/tweet-scraper)

Get the full spoken text of any YouTube video, formatted like something a person would read: real paragraphs, grouped under the video's own chapters, with a timestamp on every line. Works on single videos or an entire channel, and you only pay for the transcripts you receive.

**Pay per delivered transcript. Videos with no captions, private videos and failed videos are
never charged.**

### What it does

Most transcription output is a wall of two-word fragments. This one returns readable paragraphs, broken on pauses and sentence endings and grouped under the video's own chapter markers, so each block already carries its topic and its timestamp. Videos without chapters come back in the same shape, so downstream code only ever handles one format.

This listing is preset to return **TEXT**. Switch `formats` for any of text, SRT or VTT.

### Paste anything YouTube gives you

One `urls` list accepts **video, playlist and channel URLs mixed together** — no separate
actor, no pre-processing. A channel URL expands to its videos, a playlist to its entries, and
`maxVideosPerSource` caps how deep each one goes.

### What you get per video

| Field | What it is |
|---|---|
| `plainText` | the full transcript as clean paragraphs, not caption fragments |
| `segments` | the timed segments behind that text |
| `chapters` | chapter-grouped segments — blocks a model can chunk. Off by default: set `includeChapters` to true, and the video has to have chapters |
| `srt` / `vtt` | real subtitle files, when you ask for those formats |
| `language`, `availableLanguages`, `isAutoGenerated` | which track you got, what else exists, and whether it was machine-made |
| `videoId`, `title`, `channel`, `durationSeconds`, `uploadDate`, `url` | the video's own metadata |

Caption fragments are stitched into sentences and paragraphs before delivery, so the output is
readable — and chunk-ready for a RAG or search pipeline — instead of a wall of 2-second lines.

### Freshness: re-run without re-paying

Set `onlyNewVideos` and a re-run skips every video already delivered under the same
`stateLabel`. Point it at a channel, schedule it weekly, and you pay only for what is new.

### Languages

Ask for specific `languages`, or set `allLanguages` to take every transcript track a video
publishes. Where a video has no captions at all, the row comes back labelled — never guessed,
never silently dropped.

### Pricing

**$3.00 per 1,000 transcripts.** One price, every plan tier, no hidden multiplier.

Free: videos with no captions, private or removed videos, and failed fetches. You are charged
only for a transcript you actually receive. Cap any run with `maxVideosPerSource`, or set a
maximum cost on the run itself.

### Use it from anywhere

```bash
curl -X POST "https://api.apify.com/v2/acts/vonsensey~youtube-video-transcription/run-sync-get-dataset-items?token=YOUR_TOKEN" \
  -H 'Content-Type: application/json' \
  -d '{"urls":["https://www.youtube.com/watch?v=jNQXAC9IVRw"]}'
```

```python
from apify_client import ApifyClient
client = ApifyClient("YOUR_TOKEN")
run = client.actor("vonsensey/youtube-video-transcription").call(run_input={
    "urls": ["https://www.youtube.com/watch?v=jNQXAC9IVRw"],
})
for row in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(row.get("title"), "-", (row.get("plainText") or "")[:120])
```

Also available through the Apify integrations for **n8n, Make and Zapier**, and via **MCP** for
AI agents.

### Legal

Unofficial tool. **Not affiliated with, endorsed by, or sponsored by YouTube or Google.**
"YouTube" is a trademark of Google LLC. Reads publicly available caption tracks only — no
login, no DRM circumvention, no private or unlisted content.

### The rest of the family

- [YouTube Transcript Scraper — RAG-Ready](https://apify.com/vonsensey/youtube-transcript-scraper)
- [YouTube Transcript API - Bulk Video Transcripts, Fast](https://apify.com/vonsensey/youtube-transcript-api)
- [YouTube Subtitle Downloader - SRT & VTT Subtitle Files](https://apify.com/vonsensey/youtube-subtitle-downloader)
- [YouTube Shorts Transcript Scraper - Bulk Shorts to Text](https://apify.com/vonsensey/youtube-shorts-transcript-scraper)
- [YouTube Channel Transcript Scraper - Whole Channel to Text](https://apify.com/vonsensey/youtube-channel-transcript-scraper)
- [YouTube Playlist Transcript Scraper - Whole Playlist to Text](https://apify.com/vonsensey/youtube-playlist-transcript-scraper)
- [YouTube Closed Captions Extractor - Every Caption Track](https://apify.com/vonsensey/youtube-closed-captions-extractor)
- [YouTube Transcript RAG Dataset Builder - Chunk-Ready Text](https://apify.com/vonsensey/youtube-transcript-rag-dataset)

Issues and requests go in the **Issues** tab — first response within 24 hours.

### Use it from n8n, MCP, the API or a schedule

Built to be called by a workflow, not only from the Store form. The Actor is `vonsensey/youtube-video-transcription`; every snippet below sends `{}`, which runs the defaults shown on the form — replace it with your own input.

#### n8n

Install the **Apify** community node (`@apify/n8n-nodes-apify` under *Settings → Community Nodes*, or search "Apify" on n8n Cloud). Add **Apify → Run Actor** with Actor `vonsensey/youtube-video-transcription` and your input JSON, then **Apify → Get Dataset Items** on the run's `defaultDatasetId` and pipe the rows anywhere. For scheduled runs, the **On new Apify Event** trigger fires when a run of this Actor finishes.

#### MCP (Claude, Cursor, VS Code, any MCP client)

```json
{
  "mcpServers": {
    "apify": {
      "url": "https://mcp.apify.com?tools=vonsensey/youtube-video-transcription",
      "headers": {
        "Authorization": "Bearer <YOUR_APIFY_TOKEN>"
      }
    }
  }
}
```

Your agent then calls `vonsensey/youtube-video-transcription` as a tool with the same input the form takes and reads the dataset back.

#### REST API (one call, rows in the response)

```bash
curl -X POST "https://api.apify.com/v2/acts/vonsensey~youtube-video-transcription/run-sync-get-dataset-items?token=$APIFY_TOKEN" \
  -H "Content-Type: application/json" -d '{}'
```

#### Python

```python
from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("vonsensey/youtube-video-transcription").call(run_input={})
for row in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(row)
```

#### JavaScript

```js
import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('vonsensey/youtube-video-transcription').call({});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
```

#### Make, Zapier, LangChain, CrewAI

The Apify app in **Make** and **Zapier** has a *Run an Actor* module: pick `vonsensey/youtube-video-transcription`. In **LangChain** and **CrewAI** the Apify tool wrappers take the same Actor id. A daily **schedule** needs nothing but the Console: *Schedules → Create → this Actor → cron*, and the dataset fills on its own.

> **Run it without configuring anything** — [Turn a YouTube video into text](https://apify.com/vonsensey/youtube-video-transcription/examples/youtube-video-to-text), a ready-made example you can start as-is or copy.

### Use cases

- **Build a RAG corpus.** Turn a channel, playlist or URL list into chunk-ready paragraphs with the source video and timestamp on every row — ready to embed.
- **Repurpose long video.** Pull the spoken text of a talk or podcast and turn it into a post, a newsletter or a clip list without watching it.
- **Search what was said.** Make a back catalogue greppable: find every mention of a product, a name or a claim across hundreds of videos.
- **Keep a corpus fresh.** Schedule it and re-runs skip videos already delivered, so you pay for new material only.

### Run it on a schedule

A one-off pull answers a question; a schedule answers it every day without you. Open **Schedules** in the Apify Console, point a cron at this Actor, and the dataset keeps filling on its own — no server, no cron box, no babysitting. Everything here is built to be re-run: you are billed per transcript delivered, and the FAQ below says exactly what a scheduled run that finds nothing new costs.

### FAQ

#### Do I need a YouTube API key?

No. No key, no login, no OAuth and no YouTube quota to manage — you give it a URL and it returns text.

#### Can I export transcripts to CSV, JSON or Excel?

Yes. Every run writes a dataset you can export in one click from the Console, or pull straight from the API in JSON, CSV, XLSX or JSONL.

#### What happens to a video with no captions?

It comes back as a free row that says so, rather than failing the run. You are never charged for a video that returned no transcript.

#### Can I transcribe a whole channel or playlist at once?

Yes — pass the channel or playlist URL and it walks the uploads for you. There are dedicated Actors in this suite for both.

***

Something wrong, or a field you need that is missing? Open an issue on the **Issues** tab — it is read and it gets fixed. If this saved you time, a rating on the Store page helps the next person find it.

# Actor input Schema

## `urls` (type: `array`):

Mixed list of YouTube URLs. Individual videos, playlists, channels, and @handles are all accepted — each playlist or channel URL is expanded into its videos automatically.

## `languages` (type: `array`):

Ordered language preference for the transcript. BCP-47 tags or prefixes are accepted (e.g. "en" matches en, en-US, en-GB; "pt-BR" matches only Brazilian Portuguese). The first available match wins. Leave empty to take the video's default caption track.

## `allLanguages` (type: `boolean`):

Return every available caption track instead of just the best language match. Warning: this returns one dataset item per caption track, and each track is billed as one transcript.

## `includeShorts` (type: `boolean`):

When expanding channel URLs, also include Shorts. By default channel expansion covers long-form uploads only.

## `maxVideosPerSource` (type: `integer`):

Maximum number of videos to take from each playlist or channel URL, applied per source URL independently. Direct video URLs are not affected.

## `includeChapters` (type: `boolean`):

Extract chapter markers and upload date for each video. Disabling this saves one request per video and leaves uploadDate null.

## `formats` (type: `array`):

Additional rendered formats to include on each dataset item. Structured JSON segments are always included.

## `onlyNewVideos` (type: `boolean`):

Skip videos already delivered by previous runs sharing the same state label. Useful for scheduled re-runs that should only pick up new uploads.

## `stateLabel` (type: `string`):

Names the persistent delivered-state store used by "Only new videos". Runs sharing a label share the same delivered-video memory.

## `proxyConfiguration` (type: `object`):

Proxy settings for YouTube requests. Residential is the default because YouTube blocks most datacenter IPs for caption fetching (measured ~21% success on datacenter vs ~95% on residential). Keep residential unless you have your own proxy URLs. Optional — proxy is handled for you and its cost is included in the price. Only set this if you have a specific reason.

## `spikeConfig` (type: `object`):

Internal: proxy-economics spike configuration. Leave empty.

## Actor input object example

```json
{
  "urls": [
    "https://www.youtube.com/watch?v=jNQXAC9IVRw"
  ],
  "allLanguages": false,
  "includeShorts": false,
  "maxVideosPerSource": 25,
  "includeChapters": true,
  "formats": [
    "text"
  ],
  "onlyNewVideos": false,
  "stateLabel": "default",
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# Actor output Schema

## `transcripts` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "urls": [
        "https://www.youtube.com/watch?v=jNQXAC9IVRw"
    ],
    "formats": [
        "text"
    ],
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("vonsensey/youtube-video-transcription").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "urls": ["https://www.youtube.com/watch?v=jNQXAC9IVRw"],
    "formats": ["text"],
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("vonsensey/youtube-video-transcription").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "urls": [
    "https://www.youtube.com/watch?v=jNQXAC9IVRw"
  ],
  "formats": [
    "text"
  ],
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call vonsensey/youtube-video-transcription --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,vonsensey/youtube-video-transcription"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/dfDM0OnMqbY7pfqhl/builds/SysTbUyVrimDaC4rW/openapi.json
