# TikTok Transcript Scraper (`logical_scrapers/tiktok-transcript-scraper`) Actor

Get the transcript of public TikTok videos as plain text and timed segments, with its language, whether it is auto-generated, the other subtitle languages, the author and view, like, comment and share counts.

- **URL**: https://apify.com/logical\_scrapers/tiktok-transcript-scraper.md
- **Developed by:** [Goldmine](https://apify.com/logical_scrapers) (community)
- **Categories:** Social media, Videos, AI
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.55 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## 📝 TikTok Transcript Scraper — Get the Transcript of Any TikTok Video as Text

![TikTok Transcript Scraper](https://i.ibb.co/Zz2QLg90/Screenshot-2026-09-17-at-7-36-39-PM.png)

This **TikTok transcript scraper** turns the subtitles of public videos on **[TikTok](https://www.tiktok.com/)** into text you can search, analyse and reuse. Give it video URLs or share links and it returns each video's **TikTok transcript** as plain text and as timed segments, with the language, whether the captions were auto-generated or written by the creator, the other languages TikTok offers, and the video's author, publish date, duration and view, like, comment and share counts. People use it to **scrape TikTok subtitles** for content research, to repurpose videos into blog posts and captions, to monitor what creators and brands say, and to feed spoken content into AI pipelines.

Paste video links, pick a language, and get one dataset item per video, ready to export to JSON, CSV or Excel. No TikTok account, cookies, API key or proxy is needed.

***

### 🚀 Key Features

- 🔎 **Any video link** — full video URLs, `vm.tiktok.com` and `vt.tiktok.com` share links and older `m.tiktok.com/v/…` links
- 🧾 **Transcript as text and segments** — the full transcript in one field, plus every caption line with its start and end time in seconds
- 🌍 **Languages** — the language spoken in the video by default, a language of your choice (for example English translations of Spanish videos), or every track TikTok has
- 🏷️ **Caption source** — whether the transcript was generated automatically, written by the creator, or machine-translated
- 📈 **Video details on every item** — author, caption, hashtags, publish time, duration, views, likes, comments, shares and saves
- 📑 **One item per video** — each URL gives at most one result, so the count you pay for is the number of videos with a transcript
- 🧹 **No charge for videos without subtitles** — music-only videos, photo posts, removed and private videos are listed in the run's key-value store instead of the dataset
- 🛡️ **No login and no proxy** — reads only what TikTok shows logged-out visitors
- 📤 **Multiple export formats** — JSON, CSV, Excel, XML via the Apify dataset
- 🔁 **Schedulable runs** — run it on a schedule over the videos you track

***

### 👥 Who Is This Actor For?

- 📣 **Marketers and social media teams** — reading what competitors and creators actually say in their videos, not just their captions
- ✍️ **Content creators and agencies** — turning TikTok videos into blog posts, newsletters, subtitles and scripts
- 🧠 **Researchers and analysts** — building datasets of spoken TikTok content in several languages
- 🔍 **Brand and compliance monitoring** — checking videos that mention a brand or product for what was said
- 🤖 **AI builders** — feeding video transcripts into summarisation, search, classification or RAG pipelines

### 💡 Common Use Cases

- Get the transcript of a TikTok video as text
- Scrape TikTok subtitles for a list of videos into a spreadsheet
- Download TikTok captions with timestamps to reuse as subtitles
- Translate TikTok videos to English using TikTok's own translated captions
- Summarise a batch of TikTok videos with an LLM from their transcripts
- Search what creators said about a product across many videos

***

### 🔗 Supported URL Types

| URL Type | Example |
| -------- | ------- |
| **Video URL** | `https://www.tiktok.com/@plantyou/video/7463223840278187269` |
| **Share link (vm)** | `https://vm.tiktok.com/ZM…/` |
| **Share link (vt)** | `https://vt.tiktok.com/ZS…/` |
| **Old mobile link** | `https://m.tiktok.com/v/7484309187241987350.html` |

***

### 📥 Input

| Field | Type | Description | Default |
| ----- | ---- | ----------- | ------- |
| `startUrls` | Array | TikTok videos to get transcripts for: video URLs or share links. One transcript per video. | Required |
| `language` | String | Which subtitle track to return. `original` returns the language spoken in the video. A language code such as `en` or `es` returns that language when TikTok has it (often a machine translation) and falls back to the original otherwise. `all` returns the original in the main fields plus every track in `tracks`. | `original` |
| `includeVtt` | Boolean | Also return the raw WebVTT subtitle file in the `vtt` field. | `false` |
| `includeVideosWithoutTranscript` | Boolean | Also save a result for videos that have no subtitles, with `transcript` set to `null` and a `note` saying why. These are charged like any other result. When off, such videos are only listed in the `VIDEOS_WITHOUT_TRANSCRIPT` record of the key-value store, free of charge. | `false` |
| `proxyConfiguration` | Object | Proxy settings. TikTok serves video pages without a proxy, so none is used by default. If videos come back blocked, switch to Apify datacenter proxies. | `{ "useApifyProxy": false }` |

Because each URL is one video, there is no `maxItems` limit: a run returns at most one item per start URL.

#### Example Input

```json
{
  "startUrls": [
    { "url": "https://www.tiktok.com/@ethan_steele_/video/7484309187241987350" },
    { "url": "https://www.tiktok.com/@ale.vegana/video/7592801339265060103" }
  ],
  "language": "original"
}
```

***

### 📤 Output

Each dataset item is one video:

| Field | Type | Description |
| ----- | ---- | ----------- |
| `videoId` | String | The video's id |
| `url` | String | Link to the video |
| `inputUrl` | String | The URL or share link you gave |
| `description` | String | The video's caption |
| `hashtags` | Array | Hashtags in the caption, without `#` |
| `authorUsername` | String | The creator's username |
| `authorNickname` | String | The creator's display name |
| `authorId` | String | The creator's user id |
| `publishedAt` | String | Publish time (ISO 8601, UTC) |
| `durationSeconds` | Number | Video length in seconds |
| `playCount` | Number | Views |
| `likeCount` | Number | Likes |
| `commentCount` | Number | Comments |
| `shareCount` | Number | Shares |
| `saveCount` | Number | Saves (bookmarks) |
| `hasTranscript` | Boolean | Whether the item carries a transcript |
| `transcript` | String | The transcript as plain text, caption lines joined |
| `transcriptSegments` | Array | Caption lines as `{ start, end, text }`, times in seconds |
| `language` | String | Language of the returned transcript, e.g. `en` |
| `languageTag` | String | TikTok's language tag, e.g. `eng-US` |
| `isAutoGenerated` | Boolean | `true` for automatic captions, `false` for captions written by the creator |
| `isTranslation` | Boolean | `true` when the transcript is a machine translation of the spoken language |
| `originalLanguage` | String | The language spoken in the video, as TikTok detected it |
| `availableLanguages` | Array | Every subtitle language TikTok offers for the video |
| `subtitleUrl` | String | Link to TikTok's subtitle file. It expires about 48 hours after the run |
| `subtitleUrlExpiresAt` | String | When `subtitleUrl` stops working (ISO 8601, UTC) |
| `vtt` | String | The raw WebVTT file, only with `includeVtt` |
| `tracks` | Array | With `language: "all"`: every track, each with its own `language`, `isAutoGenerated`, `isTranslation`, `transcript`, `transcriptSegments` and `subtitleUrl` |
| `note` | String | Set when the requested language was not available, or the video has no subtitles; otherwise `null` |
| `scrapedAt` | String | When the item was scraped (ISO 8601, UTC) |

Videos that return no transcript (no subtitles, photo posts, removed, private or unavailable videos, and links that do not lead to a video) are listed with the reason in the `VIDEOS_WITHOUT_TRANSCRIPT` record of the run's key-value store, and counted in the run's status message.

#### Example Output

A real item from a run, with the subtitle URL shortened and the segment list cut to three of its seventeen lines:

```json
{
  "videoId": "7463223840278187269",
  "url": "https://www.tiktok.com/@plantyou/video/7463223840278187269",
  "inputUrl": "https://www.tiktok.com/@plantyou/video/7463223840278187269",
  "description": "Chipotle Freezer Burritos #burritos #mealprep #foodprep #easyrecipe #plantbased ",
  "hashtags": ["burritos", "mealprep", "foodprep", "easyrecipe", "plantbased"],
  "authorUsername": "plantyou",
  "authorNickname": "Carleigh Bodrug - plantyou",
  "authorId": "6623907532170182661",
  "publishedAt": "2025-01-23T21:20:17.000Z",
  "durationSeconds": 37,
  "playCount": 3200000,
  "likeCount": 106800,
  "commentCount": 569,
  "shareCount": 16100,
  "saveCount": 95511,
  "hasTranscript": true,
  "transcript": "Here's how you're gonna fill your freezer with high protein veggie pack Chipotle black bean burritos in under an hour. This is part of my series of low effort recipes to help you eat more plants. First, add diced potatoes, bell peppers, onions and tomatoes to a sheet pan with a little bit of oil, cumin, paprika and salt. Bake that for about 45 minutes. […] Assemble and make sure to press follow if you want to eat healthier with me in 2025.",
  "transcriptSegments": [
    { "start": 0, "end": 2.82, "text": "Here's how you're gonna fill your freezer with high protein veggie" },
    { "start": 2.821, "end": 5.781, "text": "pack Chipotle black bean burritos in under an hour." },
    { "start": 5.782, "end": 8.141, "text": "This is part of my series of low effort recipes" }
  ],
  "language": "en",
  "languageTag": "eng-US",
  "isAutoGenerated": true,
  "isTranslation": false,
  "originalLanguage": "en",
  "availableLanguages": ["en"],
  "subtitleUrl": "https://v16m-webapp.tiktokcdn-us.com/e653862df561a5ad0d5333f74f438cfb/6abe9191/video/tos/useast2a/…",
  "subtitleUrlExpiresAt": "2026-10-01T17:00:01.000Z",
  "note": null,
  "scrapedAt": "2026-09-29T16:59:24.000Z"
}
```

You can export the dataset as **JSON, CSV, Excel, XML, RSS, or HTML** from the Apify Console or via the [Apify API](https://docs.apify.com/api/v2).

***

### 💰 Pricing

This Actor is **pay per event** (you pay per result): each video saved with a transcript costs **$0.00299 on the free plan, so 1,000 videos cost $2.99**, falling to $0.00255 per video ($2.55 per 1,000) on Gold and higher Apify plans. Each run also has a $0.00005 start fee per GB of memory. Videos without subtitles are not charged unless you turn on `includeVideosWithoutTranscript`. It runs without a proxy by default, so there is no proxy traffic to pay for. A free Apify account is enough to try it.

***

### ❓ FAQ

#### What is the TikTok Transcript Scraper?

It is a TikTok transcript scraper that takes public video URLs or share links and returns each video's transcript as text and timed segments, with its language and the video's details. You can run it from the Apify Console, the API, or on a schedule.

#### Where does the transcript come from?

From the subtitles TikTok itself shows on the video: automatic captions generated from the speech, captions written by the creator, and TikTok's translations of them. The Actor does not transcribe audio, so a video without TikTok subtitles has no transcript.

#### Which videos have no transcript?

Videos without speech (music only), videos TikTok has not generated captions for, videos whose captions are turned off, and photo posts. These are not charged by default; they are listed with the reason in the `VIDEOS_WITHOUT_TRANSCRIPT` record of the run's key-value store.

#### Can I get the transcript in English?

Set `language` to `en`. For videos in another language TikTok often has an English machine translation, which the Actor returns with `isTranslation: true`. When a video has no English track you get the original-language transcript and a `note` saying so.

#### Do I need an account, cookies or an API key?

No. The Actor reads only what TikTok shows to visitors who are not logged in.

#### Do I need a proxy?

No. The default runs without a proxy. If videos come back blocked, switch the proxy setting to Apify datacenter proxies.

#### How many results can I get per run?

One item per video URL, so a run with 500 URLs returns up to 500 transcripts. There is no `maxItems` setting because each URL is a single video.

#### Why does the subtitle URL stop working?

TikTok signs subtitle links for about 48 hours (`subtitleUrlExpiresAt`). The Actor downloads the file during the run and returns the text, so the transcript itself does not expire. Turn on `includeVtt` to keep the original WebVTT file.

#### Can I schedule it or integrate it with my app?

Yes. Every Apify Actor exposes a REST API, webhooks, and integrations with Zapier, Make, Google Sheets and Slack, and can be called from an AI agent via the Apify MCP server.

#### Is it legal to scrape TikTok?

This Actor accesses only publicly available pages. You are responsible for making sure your use complies with TikTok's terms of service and the laws in your jurisdiction.

#### What if TikTok changes and the Actor breaks?

Open an issue on the Actor's **Issues** tab with a video URL that fails, so the problem can be reproduced.

***

### 🧩 Part of the TikTok Scraper Suite

Goldmine's TikTok Scraper Suite covers search results, posts, comments, transcripts and video files. Also try:

- [TikTok Post Scraper](https://apify.com/logical_scrapers/tiktok-post-scraper) — views, likes, shares, hashtags and author for TikTok posts
- [TikTok Comments Scraper](https://apify.com/logical_scrapers/tiktok-comments-api) — comments and reply threads from public TikTok videos
- [TikTok Search Scraper](https://apify.com/logical_scrapers/tiktok-search-scraper) — find TikTok videos by keyword, with full stats per result
- [TikTok Video Downloader](https://apify.com/logical_scrapers/tiktok-video-downloader) — download TikTok videos as MP4 files
- [Scrape TikTok Post Details](https://apify.com/logical_scrapers/tiktok-realtime-post-details-scraper) — likes, plays, comments, cover and video links for TikTok posts

See every Goldmine Actor at [apify.com/logical\_scrapers](https://apify.com/logical_scrapers).

***

### 📷 Image Credit

Image credit: [tiktok.com](https://www.tiktok.com/)

***

### 📬 Contact & Support

- **Issues & feature requests**: use the Issues tab on this Actor's page. Include a video URL that fails so it can be reproduced.
- **Email**: `coredev.dan@gmail.com`
- **If this Actor saved you time, please leave a ⭐ rating on the Apify Store.** It helps us keep it maintained.

# Actor input Schema

## `startUrls` (type: `array`):

TikTok videos to get transcripts for: video URLs such as https://www.tiktok.com/@user/video/1234567890, or share links such as https://vm.tiktok.com/ZM.../ and https://vt.tiktok.com/ZS.../. One transcript per video.

## `language` (type: `string`):

Which subtitle track to return. "original" returns the language spoken in the video. A language code such as "en" or "es" returns that language when TikTok has it (often a machine translation) and falls back to the original otherwise. "all" returns the original in the main fields plus every track in "tracks".

## `includeVtt` (type: `boolean`):

Also return the raw WebVTT subtitle file in the "vtt" field.

## `includeVideosWithoutTranscript` (type: `boolean`):

Also save a result for videos that have no subtitles, with "transcript" set to null and a "note" saying why. These are charged like any other result. When off, such videos are only listed in the VIDEOS\_WITHOUT\_TRANSCRIPT record of the key-value store, free of charge.

## `proxyConfiguration` (type: `object`):

Proxy settings. TikTok serves video pages without a proxy, so none is used by default. If videos come back blocked, switch to Apify datacenter proxies.

## Actor input object example

```json
{
  "startUrls": [
    {
      "url": "https://www.tiktok.com/@ethan_steele_/video/7484309187241987350"
    },
    {
      "url": "https://www.tiktok.com/@ale.vegana/video/7592801339265060103"
    }
  ],
  "language": "original",
  "includeVtt": false,
  "includeVideosWithoutTranscript": false,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# Actor output Schema

## `transcripts` (type: `string`):

No description

## `videosWithoutTranscript` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        {
            "url": "https://www.tiktok.com/@ethan_steele_/video/7484309187241987350"
        },
        {
            "url": "https://www.tiktok.com/@ale.vegana/video/7592801339265060103"
        }
    ],
    "language": "original"
};

// Run the Actor and wait for it to finish
const run = await client.actor("logical_scrapers/tiktok-transcript-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": [
        { "url": "https://www.tiktok.com/@ethan_steele_/video/7484309187241987350" },
        { "url": "https://www.tiktok.com/@ale.vegana/video/7592801339265060103" },
    ],
    "language": "original",
}

# Run the Actor and wait for it to finish
run = client.actor("logical_scrapers/tiktok-transcript-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    {
      "url": "https://www.tiktok.com/@ethan_steele_/video/7484309187241987350"
    },
    {
      "url": "https://www.tiktok.com/@ale.vegana/video/7592801339265060103"
    }
  ],
  "language": "original"
}' |
apify call logical_scrapers/tiktok-transcript-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,logical_scrapers/tiktok-transcript-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/yYTKBDpIKGSptrvM3/builds/BChQpf2JJHkDEtKXL/openapi.json
