Instagram Transcript Scraper
Pricing
from $2.49 / 1,000 results
Instagram Transcript Scraper
Instagram Transcript Scraper SD - Instagram Transcript Scraper is a social media data tool that extracts spoken-word transcripts of Instagram reels and videos with timestamps and detected language, using speech recognition - Instagram reel to text converter.
Pricing
from $2.49 / 1,000 results
Rating
0.0
(0)
Developer
Neuro Scraper
Maintained by CommunityActor stats
0
Bookmarked
11
Total users
0
Monthly active users
2 days ago
Last modified
Share
Instagram Transcript Scraper — reels to text with timestamps
The Instagram Transcript Scraper turns public Instagram reels and videos into text: it downloads the audio track and transcribes it with speech recognition, returning the full transcript, timestamped segments and the detected language.
It runs on Apify, needs no Instagram login and no cookies, and exports to JSON, CSV, Excel or an API call. Every session key it needs (tokens, query ids, cursors) is read from Instagram on each run, so it keeps working when Instagram ships new code.
What the Instagram Transcript Scraper does
- Uses the audio-only track Instagram publishes, which keeps downloads small.
- Transcribes with a Whisper speech model built into the Actor.
- Detects the language or uses the one you set.
- Marks music-only reels with hasSpeech=false instead of inventing words.
Who uses the Instagram Transcript Scraper
- Creators repurposing reels into blog posts and captions.
- Agencies analysing what competitors say on video.
- Researchers studying spoken content.
How the Instagram Transcript Scraper works
Instagram serves every public page to logged-out visitors together with the queries that fill it.
The Actor loads the page the way a browser does, reads the session values Instagram puts into it, and then asks Instagram's own GraphQL endpoint for the next page with the cursor the previous page returned.
Query ids that a page does not preload are read from the JavaScript bundles that page lists. Large files from Instagram's CDN (images, videos) are fetched directly with their signed links, so proxy traffic stays small.
Nothing is guessed: if Instagram stops publishing a value the Instagram Transcript Scraper needs, the run log names it and the item is skipped instead of being filled with a stale value.
What is new in the Instagram Transcript Scraper (October 2026)
These features were added on 2026-10-03 after comparing the Instagram Transcript Scraper with the ten most used Actors of its kind on the Apify Store. Defaults keep the earlier behaviour.
wordTimestampsadds every word with its start and end time (100 words for a 38.6 s NASA reel).translateToEnglishadds an English translation from the same speech model (measured: a Spanish BBC Mundo reel, detected es 0.998, in 42 to 43 seconds).includeSubtitlesadds SRT and WebVTT subtitles;includeSegments(on by default) can drop the segment list.maxDurationSeconds,skipNoSpeech,onlyPostsNewerThanandskipPinnedPostschoose which videos are transcribed.- New fields: id, commentsCount, hashtags, mentions, ownerFullName, productType, videoUrl and audioUrl.
Input of the Instagram Transcript Scraper
| Field | Type | Default | What it does |
|---|---|---|---|
videoUrls | array | ["https://www.instagram.com/reel/Dd1gQBsCWE8/"] | Links to Instagram reels or video posts. |
usernames | array | [] | One per line: a username (nasa), an @handle (@nasa) or a profile URL (https://www.instagram.com/nasa/). |
maxPerProfile | integer | 5 | For usernames only. |
language | string | — | Two-letter code such as en, es, hi. Empty = detect automatically. |
maxDurationSeconds | integer | 0 | Videos longer than this are not transcribed. |
includeSegments | boolean | True | Adds the start / end time of every sentence. |
wordTimestamps | boolean | False | Adds a words list with the start / end time of every word (slower). |
translateToEnglish | boolean | False | Adds an English translation made by the same speech model (one more pass per video). |
includeSubtitles | boolean | False | Adds srt and vtt subtitle text built from the segments. |
skipNoSpeech | boolean | False | Music-only videos are not returned. |
onlyPostsNewerThan | string | — | For usernames only (2026-09-01, an ISO time or "7 days"). |
skipPinnedPosts | boolean | False | For usernames only: leave out reels pinned to the top of the Reels tab. |
maxItems | integer | 0 | Stops the whole run after this many dataset rows. 0 means only the per-target limits apply. |
proxyConfiguration | object | {"useApifyProxy": true, "apifyProxyGroups": ["RESIDENTIAL"]} | Instagram limits logged-out visitors per IP address. Residential proxies (the default) give the most reliable results; the Actor switches to a fresh IP when Ins |
Example input for the Instagram Transcript Scraper:
{"videoUrls": ["https://www.instagram.com/reel/Dd1gQBsCWE8/"],"maxPerProfile": 5}
Output of the Instagram Transcript Scraper
Each dataset row is one transcribed video. These are the fields the Instagram Transcript Scraper produced in its test runs:
| Field | Meaning |
|---|---|
caption | Caption text |
hasSpeech | False when the audio has no speech (music only) |
inputUrl | inputUrl |
language | Detected language |
languageProbability | languageProbability |
likesCount | Likes (or reactions) count |
ownerUsername | Account that posted it |
scrapedAt | When this row was read |
segments | Transcript segments with start/end seconds |
shortCode | Instagram short code |
timestamp | Publication time (UTC, ISO 8601) |
transcript | Full transcript text |
url | Public link of the item |
videoDurationSec | Video length in seconds |
videoPlayCount | Video plays |
audioUrl | Audio-only track link (signed CDN link, expires) |
commentsCount | Comments count |
hashtags | Hashtags found in the caption |
id | Platform id of the item |
mentions | @accounts mentioned in the caption |
ownerFullName | Owner display name |
productType | feed, carousel_container or clips (reel) |
videoUrl | Video file link (signed CDN link, expires) |
srt | SRT subtitles |
translation | English translation of the speech |
vtt | WebVTT subtitles |
words | Every word with start / end seconds |
Fields that Instagram does not publish for an item are left empty rather than guessed. Download the dataset as JSON, CSV, Excel, XML or HTML, or read it through the Apify API.
Example row from a Instagram Transcript Scraper test run
A real row from the local test run (long values shortened):
{"url": "https://www.instagram.com/reel/Dd_irlvg4ar/","shortCode": "Dd_irlvg4ar","ownerUsername": "natgeo","caption": "Presented by @Rolex. A simple branch becomes a fishing rod in the hands of these resour...","timestamp": "2026-10-02T13:01:26.000Z","videoDurationSec": 29.95,"videoPlayCount": 2472781,"likesCount": 83187,"inputUrl": "natgeo","language": "en","languageProbability": 0.61,"transcript": "","segments": [],"hasSpeech": false,"scrapedAt": "2026-10-02T18:48:08.000Z"}
Data quality of the Instagram Transcript Scraper
- Every value in a Instagram Transcript Scraper row is read from Instagram during your run; nothing is filled from a cache or an old snapshot.
- Rows are de-duplicated by id inside a run, so the same transcribed video never appears twice in one dataset.
- Counts are numbers you can sort and sum (not "1.2K" text) wherever Instagram publishes the exact number.
- Dates are UTC in ISO 8601 format, which Excel, Google Sheets, Python and SQL all read directly.
Instagram Transcript Scraper test results
These numbers come from live test runs of the Instagram Transcript Scraper on 2026-10-03 (home connection, no proxy). Runs on the Apify platform use residential proxies and can be faster or slower.
| Test | Rows | Time |
|---|---|---|
| prefill input (regression) | 1 | 15 s |
| transcribe one reel | 1 | 13 s |
| latest 2 reels of a profile | 2 | 15 s |
| NEW words + subtitles + English translation (Spanish reel) | 1 | 39 s |
Instagram Transcript Scraper limits, stated plainly
- Speech recognition is not perfect: names and numbers can be misheard. Instagram publishes no caption track logged out, so every transcript comes from the audio.
- Measured locally: one 38.6 s reel transcribed in about 22 seconds including download.
- Music-only reels return an empty transcript with hasSpeech=false.
How to run the Instagram Transcript Scraper
- Open the Instagram Transcript Scraper in Apify Console and fill in the input fields described above.
- Keep the default residential proxy unless you have a reason to change it; Instagram limits logged-out visitors per IP address.
- Start the run, watch the log, and download the dataset when the status says Done.
- To automate it, call the Instagram Transcript Scraper from the Apify API, schedule it, or connect it to Make, Zapier, n8n or Google Sheets through an integration or webhook.
Python example with the Apify API client:
from apify_client import ApifyClientclient = ApifyClient("<YOUR_APIFY_TOKEN>")run_input = {"videoUrls": ["https://www.instagram.com/reel/Dd1gQBsCWE8/"],"maxPerProfile": 5}run = client.actor("neuro-scraper/instagram-video-scraper-and-downloader-pro").call(run_input=run_input)for item in client.dataset(run["defaultDatasetId"]).iterate_items():print(item)
Instagram Transcript Scraper pricing
The Instagram Transcript Scraper uses pay-per-event pricing: you pay per dataset row (event "result"), plus the Apify platform usage of your run. Free-plan users get up to 30 rows per run, which is enough to test the Instagram Transcript Scraper on real data.
Set a maximum cost per run in Apify and the Instagram Transcript Scraper stops cleanly when it is reached.
Tips for large runs of the Instagram Transcript Scraper
- Start with a small limit to check the Instagram Transcript Scraper output, then raise it; the first rows arrive within seconds.
- Split very large jobs into several inputs and run them in parallel tasks; each run gets its own proxy sessions.
- Set a maximum cost per run in Apify; the Instagram Transcript Scraper stops cleanly when the budget is reached and keeps every row it already saved.
- Schedule repeat runs instead of one huge run when you track the same targets over time.
- Keep the residential proxy: datacenter IPs are refused by Instagram far more often for logged-out visitors.
Instagram Transcript Scraper compared with copying by hand
| Instagram Transcript Scraper | Manual work | |
|---|---|---|
| Speed | Seconds per page of results | Minutes per item |
| Fields | Every public field, same order every time | Whatever you copy |
| Format | JSON, CSV, Excel, API | Spreadsheet you build yourself |
| Repeatable | Schedule it daily | Start again each time |
| Errors | Logged per item | Unnoticed typos |
Using the Instagram Transcript Scraper in a workflow
- Send every new row to Google Sheets or Airtable with the Apify integration and keep a live table.
- Trigger a Make, Zapier or n8n scenario with a webhook when the run finishes.
- Load the JSON export into Python or BigQuery for analysis and dashboards.
- Use the Apify API: start the Instagram Transcript Scraper with a POST request and read the dataset items endpoint when the run finishes.
- Chain it: pass the Instagram Transcript Scraper output to another Apify Actor (for example a downloader or a transcript Actor) in an Apify task.
- Export straight to Excel from the dataset tab when you only need a one-off report.
Troubleshooting the Instagram Transcript Scraper
- The run ends with no rows: open the log. The Instagram Transcript Scraper writes one line per target saying what Instagram returned, for example a login wall or a missing item.
- Fewer rows than expected: the target may have fewer public items, or Instagram may stop serving more pages to logged-out visitors. The log names the reason.
- Slow run: large limits mean many pages; the Instagram Transcript Scraper waits briefly between pages so Instagram keeps answering. Split big jobs into several runs.
- Proxy errors: keep the RESIDENTIAL group. The Instagram Transcript Scraper already switches to a fresh IP when one is blocked.
- Free plan stops at 30 rows: that is the free-plan cap of the Instagram Transcript Scraper; any paid Apify plan removes it.
What changed in this version of the Instagram Transcript Scraper
- Rebuilt in October 2026 on Instagram's logged-out web endpoints; every token, query id and cursor is read at run time.
- Pay-per-event pricing (one event per row) replaces the old monthly rental; free-plan users can test with 30 rows per run.
- Residential proxy by default, with automatic switch to a fresh IP when Instagram blocks one.
- October 2026 feature round: new filters and fields matching the most used Instagram Actors on the Store, every one measured in a live test; existing inputs and fields are unchanged.
Is it legal to use the Instagram Transcript Scraper?
The Instagram Transcript Scraper reads only data that Instagram shows to any logged-out visitor. It never logs in and never touches private content.
Personal data may still be in public posts, so check GDPR and the other laws that apply to you before you store or reuse it, and respect Instagram's terms.
Related scrapers
| Actor | Link |
|---|---|
| Instagram Scraper | neuro-scraper/instagram-video-scraper |
| Instagram Profile Scraper | neuro-scraper/instagram-reels-scraper-downloader-pro |
| Instagram Posts Scraper | neuro-scraper/instagram-posts-scraper |
| Instagram Reels Scraper | neuro-scraper/instagram-reels-scraper |
| Instagram Hashtag Scraper | neuro-scraper/instagram-reeel-scraper |
| Instagram Comments Scraper | neuro-scraper/instagram-reels-scraper-and-downloader-premium |
| Instagram Search Scraper | neuro-scraper/instagram-video-scraper-advanced |
| Instagram Video Downloader | neuro-scraper/instagram-video-downloader |
| Instagram Reels Downloader | neuro-scraper/instagram-reels-downloader |
| Instagram Image Downloader | neuro-scraper/instagram-video-scraper-and-downloader |
| Instagram Mentions Scraper | neuro-scraper/instagram-video-scraper-and-downloader-premium |
| Instagram Keyword Scraper | neuro-scraper/instagram-video-scraper-and-downloader-fastest |
Leave a review
If the Instagram Transcript Scraper saved you time, please leave a star rating and a short review on its Apify page. Reviews help other users find the Instagram Transcript Scraper and tell us what to improve next.
Frequently asked questions
Does the Instagram Transcript Scraper invent text for music videos?
No. Segments the model rates as probably not speech are dropped, and known one-word artefacts are removed, so music-only reels return an empty transcript.
Does the Instagram Transcript Scraper need my Instagram login?
No. The Instagram Transcript Scraper works logged out and never asks for your password or cookies. It reads only what Instagram shows to any visitor.
Do I need my own proxy for the Instagram Transcript Scraper?
No. The Instagram Transcript Scraper uses Apify residential proxies by default and switches to a fresh IP when Instagram blocks one. You can choose another proxy in the input.
What formats can I export from the Instagram Transcript Scraper?
JSON, CSV, Excel, XML, HTML and RSS from the dataset tab, or any format through the Apify API, webhooks and integrations.
How fast is the Instagram Transcript Scraper?
In the test runs above the Instagram Transcript Scraper took between 13 and 39 seconds per run. Speed depends on how many items you ask for and on the proxy.
Why did the Instagram Transcript Scraper return fewer rows than I asked for?
The Instagram Transcript Scraper stops when Instagram has nothing more to show logged out, when your limit is reached, or when the free-plan cap of 30 rows applies. The run log always says which one happened.
Can I schedule the Instagram Transcript Scraper?
Yes. Create a schedule in Apify Console to run the Instagram Transcript Scraper daily or hourly and send the results to Google Sheets, Make, Zapier or n8n through an integration.
Can the Instagram Transcript Scraper read private accounts or private groups?
No. The Instagram Transcript Scraper reads public data only. Private content needs a login, and the Instagram Transcript Scraper never logs in.
What happens when Instagram changes its website?
The Instagram Transcript Scraper reads every key it needs from Instagram at run time, so routine changes are picked up automatically. If something can no longer be read, the log says so instead of returning wrong data.
Can I use the Instagram Transcript Scraper with Make, Zapier or n8n?
Yes. Add the Instagram Transcript Scraper as an Apify step in Make, Zapier or n8n, or send a webhook when the run finishes and fetch the dataset items from the Apify API.
Does the Instagram Transcript Scraper work in every country?
The Instagram Transcript Scraper reads Instagram through residential proxies that rotate across countries. Content that Instagram restricts by country follows those rules.
How do I collect only new items with the Instagram Transcript Scraper?
Schedule the Instagram Transcript Scraper and keep the item ids from earlier runs. Every row has a stable id, so a simple de-duplication against your previous export keeps only the new ones.
Can the Instagram Transcript Scraper process many targets in one run?
Yes. Put as many targets as you like in the input; the Instagram Transcript Scraper handles them one after another and a failing target never stops the others.
Can the Instagram Transcript Scraper translate transcripts?
Yes, into English. Turn on translateToEnglish: the same Whisper model translates the speech, which takes about one more pass per video.
Does the Instagram Transcript Scraper make subtitle files?
Yes. Turn on includeSubtitles and each row carries SRT and WebVTT text built from the timestamped segments.