Zoom Video Transcript
Pricing
from $0.423 / transcript
Zoom Video Transcript
Zoom Video Transcript turns one public Zoom Clip into structured meeting text for follow-ups, decisions, and searchable notes. It returns detected language, ordered timestamped segments, clip metadata, and optional translation into 133 languages. One transcript starts at $0.423.
Pricing
from $0.423 / transcript
Rating
0.0
(0)
Developer
TrueFetch
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 hours ago
Last modified
Categories
Share
Affiliate disclosure: Apify links on this page may include referral parameters. If you sign up through one of these links, TrueFetch may earn a commission from Apify at no extra cost to you. Pricing, features, and Actor access are unaffected.
Zoom Video Transcript converts one public Zoom Clip into detected-language text, timestamped segments, clip metadata, and an optional translated transcript.
- The required input is one public Zoom Clip share URL; meeting pages, protected recordings, uploads, clip libraries, and cross-platform links are outside this Actor's scope.
- Every completed Clip job publishes one 22-field item that keeps its spoken presentation, time intervals, owner details, and exposed viewing context together.
- The 133-language selector is optional, and translation is billed only after the complete ordered segment list is available.
- The URL precheck resolves the share link and confirms Zoom before media download and speech recognition begin.
Run a one-result test · View API The prefilled Clip is suitable for a minimal FREE-tier check. Fixed charges are $0.47 for its transcript and $0.008 to start the default 8 GB run, with usage measured separately. A finished Spanish version adds $0.15.
What does Zoom Video Transcript do?
Zoom Video Transcript turns speech from one public Zoom Clip into detected-language text and ordered HH:MM:SS.mmm segments. When exposed, clip metadata can include title, description, creator, duration, thumbnail, publishing time, and view count.
Choose translate to create another-language rendering of the Clip. Detected speech remains in transcript; the companion object keeps corresponding intervals and appears only when all segments succeed.
This Actor does not join meetings, browse a Zoom recording library, unlock passcoded Clips, process lists, identify speakers, capture live captions, download video, or write subtitle files. It accepts one public Clip share link.
How do I run Zoom Video Transcript?
- Open Zoom Video Transcript on Apify.
- Copy the share URL of one public Zoom Clip and paste it into Video URL.
- Set Translate only if a second language is useful; keep it blank for the detected speech alone.
- Click Start and wait for media preparation and speech recognition to finish.
- Review the one Dataset record for participant names, Clip visibility, and phrase timing before reuse.
Start with a Zoom Clip containing clear dialogue. The current input schema pre-fills this public share URL:
{"video_url": "https://zoom.us/clips/share/6YEG4i_qS0eiT7UH9zPeYw","translate": "spanish"}
The share link defines the entire job and yields at most one item. Its owner can later change access, add a passcode, or delete the Clip.
What data does Zoom Video Transcript return?
The Actor writes one successful result to the default Dataset. Its 22-field contract is nullable where Zoom omits a value.
| Group | Dataset fields | Meaning |
|---|---|---|
| Processing | processor, processed_at | Actor URL and UTC processing time |
| Zoom Clip | platform, title, description, thumbnail, published_at | Source identity, text, image, and publishing time |
| Creator | author, author_id, author_url | Creator details exposed for the media |
| Media | duration, audio_title, audio_artist | Source-reported length and audio attribution |
| Engagement | view_count, like_count, shares_count, dislike_count, comment_count | Clip activity values present in the public response; unsupported counters stay nullable |
| Labels | categories, tags | Source-provided classifications when available |
| Speech | transcript, translation | Original timestamped transcript and optional translated version |
This abbreviated response uses the same Zoom-Clip-and-Spanish scenario. Values omitted by Zoom remain null, zero, or empty; the Actor does not invent them.
{"processor": "https://apify.com/truefetch/zoom-video-transcript?fpr=aiagentapi","processed_at": "2026-07-21T12:51:46+00:00","platform": "ZoomClips","title": "Test Clip","description": null,"author": "Hans Müller","duration": 1,"view_count": 4,"transcript": {"language": "English","text": "Hello world.","segments": [{"start": "00:00:00.030","end": "00:00:00.930","text": "Hello world."}]},"translation": {"language": "Spanish","text": "Hola, mundo.","segments": [{"start": "00:00:00.030","end": "00:00:00.930","text": "Hola, mundo."}]}}
The Dataset is not a Zoom recording export. It excludes the source file, private account data, summaries, confidence scores, speaker identities, OCR, SRT/VTT artifacts, and word-level alignment.
What inputs can I configure?
| Input | Type | Required | Behavior |
|---|---|---|---|
video_url | string | Yes | One public Zoom Clip share URL. |
translate | string enum | No | One of 133 target-language values. Empty means no translation. |
video_url must be a Clip share page viewable without authentication or a passcode. The schema has no account credential, meeting ID, recording-library query, upload field, date filter, or URL collection.
Targets are lowercase language labels such as spanish, french, japanese, arabic, chinese (simplified), and chinese (traditional). Selecting one adds output beside the unchanged source transcript.
What platforms and markets does Zoom Video Transcript cover?
This Zoom-only Actor is tested with a public zoom.us/clips/share/... URL. Zoom documents that a Clip owner can share a clip and copy its link. Only links whose viewing permissions allow unauthenticated access fit this Actor's public URL workflow.
Passcode-protected or sign-in-only Clips, deleted Clips, clip libraries, Zoom meeting join links, recording-management pages, webinars, and URLs without downloadable media are not promised.
Speech language is detected from audio. Review generated text and translation for names, meeting terminology, accents, brand terms, and code-switching.
Process only Clips you are authorized to use and review the current Zoom Terms of Service for the intended workflow.
Why use Zoom Video Transcript?
| Verified capability | Practical benefit |
|---|---|
| Zoom-specific platform validation | Reject unrelated platform links before transcription begins. |
| Clip narration represented as text and intervals | Find an explanation or demo step and return to its approximate position. |
| Share-page context next to the words | Retain the available owner, title, duration, thumbnail, date, and views with the transcript. |
| Target-language segments paired to source timing | Review an asynchronous presentation across languages while preserving its original speech. |
| One Clip-shaped result | Send a selected recording into knowledge capture, product enablement, support, or an AI workflow. |
The Clip owner controls access and Zoom controls exposed metadata. Recognition can mishear participant names, acronyms, overlapping audio, or technical vocabulary; verify material used as an official record.
Who is Zoom Video Transcript for?
The Actor serves developers, distributed teams, educators, support teams, researchers, and AI builders with a specific public Clip. Typical uses include searching short demos, reviewing explanations, drafting translations, and supplying timed text downstream.
Use another solution for joining meetings, searching recordings, reading comments, handling credentials, saving the video, extracting screen text, creating legal minutes, monitoring live calls, or processing non-Zoom media.
How can I use Zoom Video Transcript through the API or MCP?
Invoke truefetch/zoom-video-transcript. A synchronous Dataset call can suit a brief Clip, but durable integrations should launch asynchronously and poll. Apify notes that the client wait may expire while the run itself continues.
curl -L "https://api.apify.com/v2/actors/truefetch~zoom-video-transcript/run-sync-get-dataset-items" \-H "Authorization: Bearer YOUR_APIFY_TOKEN" \-H "Content-Type: application/json" \-d '{"video_url": "https://zoom.us/clips/share/6YEG4i_qS0eiT7UH9zPeYw","translate": "spanish"}'
Apify's hosted MCP server can expose only this Actor to a compatible client:
{"mcpServers": {"apify": {"url": "https://mcp.apify.com?tools=truefetch/zoom-video-transcript","headers": {"Authorization": "Bearer YOUR_APIFY_TOKEN"}}}}
After connecting MCP, call truefetch/zoom-video-transcript with the Clip link and spanish, then read its Dataset item. The Apify MCP guide covers setup; the Actor API page provides client examples.
How much does Zoom Video Transcript cost?
Billing separates Actor start, measured resource use, successful transcript output, and optional completed translation. Compare this local table with the live pricing page before scheduling many Clips.
| Billed event | FREE | BRONZE | SILVER | GOLD | PLATINUM | DIAMOND | Billing unit |
|---|---|---|---|---|---|---|---|
| Actor start | $0.001 | $0.001 | $0.001 | $0.001 | $0.001 | $0.001 | One event per allocated GB, minimum one |
| Actor usage | $0.00001 | $0.00001 | $0.00001 | $0.00001 | $0.00001 | $0.00001 | Runtime, proxy, and storage usage unit |
| Transcript | $0.47000 | $0.45433 | $0.43867 | $0.42300 | $0.42300 | $0.42300 | One successful timestamped transcript |
| Translation | $0.15000 | $0.14500 | $0.14000 | $0.13500 | $0.13500 | $0.13500 | One complete translated transcript when requested |
FREE-tier fixed cost starts with eight $0.001 events for the default 8 GB allocation. One $0.47000 transcript produces $0.478 before usage; a successful $0.15000 Spanish translation raises it to $0.628. Use maxTotalChargeUsd to cap recurring Clip work.
How does Zoom Video Transcript compare with alternatives?
| Option | Best fit | Trade-off |
|---|---|---|
| Zoom Video Transcript | Converting a selected public Clip share page into timed spoken text | The caller must already have an unprotected link |
| Authorized Zoom integration | Managing meetings, cloud recordings, participants, or account resources | Requires credentials and a different workflow |
| Universal transcription Actor | Combining uploads with URLs from several services | Its input is not focused on Zoom Clip permissions |
| Human transcription | Finalizing sensitive meeting statements or formal records | Slower, but adds human interpretation and approval |
Use this Actor when a public Clip link is known and its narration is the deliverable. Video To Text covers uploads or other sites; Best Video Downloader is for obtaining media rather than structured speech.
What are the limits and troubleshooting steps?
| Symptom or limit | What it means | Next step |
|---|---|---|
| The Actor rejects the platform | The URL does not resolve as a supported public Zoom Clip | Copy the public zoom.us/clips/share/... URL |
| Video information cannot be retrieved | The Clip is protected, expired, deleted, restricted, or temporarily unavailable | Confirm the URL opens without sign-in and retry |
| The media has no usable speech | The Clip lacks audio, is silent, or is too unclear to recognize | Test a share page with audible presentation or conversation |
Metadata is empty or duration is 0 | Zoom did not return that owner, timing, date, thumbnail, or view value | Leave the documented nullable value unchanged |
| Translation is absent | One or more recognized phrases failed target-language conversion | Keep the detected transcript and retry; incomplete output is not billed |
| A synchronous API call returns HTTP 408 | The connection stopped waiting before the run finished | Poll an asynchronous run and obtain the Dataset afterward |
Time ranges mark phrases, not individual words. Remove tokens and confidential meeting material from support requests; share the run ID, a public Clip link, the expected result, and the error shown.
Frequently asked questions
Can it process a protected or sign-in-only Clip?
No. The input has no passcode, credential, or cookie field. Use a Clip that opens publicly from its share URL.
Can it crawl my Zoom recording library?
No. The production contract is one public Clip share URL per run, not account access or batch crawling.
Does it use Zoom's existing Clip transcript?
No. Recognition uses the Clip's audio stream. Existing Zoom transcripts and downloadable SRT/VTT files are not part of the Dataset item.
What happens when a Zoom Clip contains no speech?
A silent screen capture or music-only Clip can produce no speech; choose one with audible narration for transcript text.
Why are description, thumbnail, or creator ID empty?
Zoom's public share response varies by Clip. Unavailable owner, thumbnail, or description values remain nullable instead of being fabricated.
Is translation billed if one segment fails?
No. The translation event waits for a complete target-language segment sequence.
Can I use segment timestamps as finished subtitles?
They locate portions of a Clip and can seed caption editing, but they are not word-level subtitle cues or a finished caption file.
Can I schedule the same Zoom Clip URL repeatedly?
Yes. A schedule can reuse the share URL, while each run incurs its own charges and may observe changed access or metadata.
Related TrueFetch Actors
- Video To Text handles local uploads and non-Zoom links through a broader speech contract.
- Instagram Video Transcript applies a comparable one-link flow to public Instagram media.
- Best Video Downloader returns the underlying media when transcription is not the goal.
Support and last updated
- Actor page and documentation
- API reference and generated clients
- Issues
- TrueFetch community
- Direct support
Local schema, extractor behavior, metadata pricing, and official documentation links were checked on July 21, 2026. Zoom controls future Clip access and exposed fields.