Threads Post Scraper — Extract Meta Threads Posts and Replies avatar

Threads Post Scraper — Extract Meta Threads Posts and Replies

Pricing

$1.80 / 1,000 results

Go to Apify Store
Threads Post Scraper — Extract Meta Threads Posts and Replies

Threads Post Scraper — Extract Meta Threads Posts and Replies

Extract public Meta Threads posts, full discussions, replies, author metrics, and media into structured data without logins, cookies, or API keys.

Pricing

$1.80 / 1,000 results

Rating

0.0

(0)

Developer

Mikolabs

Mikolabs

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Categories

Share

🧵 Threads Post Scraper — Extract Meta Threads Posts and Replies

Extract public Meta Threads posts, full discussions, replies, author metrics, and media into structured data without logins, cookies, or API keys.

Apify Actor Platform Authentication Proxies


⚡ At a Glance

FeatureDetails
Target PlatformMeta Threads (threads.net and threads.com)
AuthenticationNone required — no cookies, passwords, or developer API tokens
Data ExtractedPost text, collapsed attachments, author profiles, verified badges, engagement (likes/replies), media (images/videos), full reply threads
Export FormatsJSON, JSONL, CSV, Excel (XLSX), XML, HTML, RSS
Residential ProxiesPre-configured and enabled by default for seamless anti-bot bypass
AutomationFull REST API, Apify Python/JS SDKs, webhooks, and cron schedules

📌 What does Threads Post Scraper do?

Threads Post Scraper enables you to extract public data from Meta Threads into clean, ready-to-use structured datasets. Simply provide one or more Threads post URLs, and the scraper automatically extracts the complete post content along with all visible replies and nested comments.

Comprehensive Data Coverage

  • 📝 Full Post Contents: Captions, body text, long-form text attachments, and collapsed snippets hidden inside mobile preview toggles.
  • 💬 Complete Discussion Threads: All visible replies and nested conversations (replies to replies) with full author attribution.
  • 👤 Author Profiles & User Metrics: Username, author profile picture (avatar), verified status (user_verified), user ID, and profile links.
  • 🖼️ Rich Media Content: High-resolution image URLs, carousel sets, video URLs, and audio availability flags.
  • 📊 Engagement Metrics: Up-to-date like counts and reply counts for both root posts and replies.
  • ⏱️ Metadata & Timestamps: Exact publication times (Unix epoch timestamps), unique post IDs, numeric PKs, shortcodes, and canonical URLs.

🎯 Why scrape Meta Threads?

Meta Threads has hundreds of millions of active users and is rapidly growing into the leading platform for real-time conversations, breaking updates, and community discourse. It provides an unmatched stream of consumer sentiment, industry discussions, and viral commentary.

Here are some of the most popular ways businesses and researchers use Threads data:

  • 📈 Brand Monitoring & Reputation Management: Monitor brand mentions, track sentiment, and identify emerging customer issues before they escalate.
  • 🔍 Market Research & Trend Discovery: Follow emerging industry topics, viral memes, and discussions in real time.
  • 🏆 Competitor Intelligence: Benchmark competitors' post frequency, engagement rates, and see what their audience discusses in comment sections.
  • 🗣️ Voice of Customer (VoC): Collect authentic user feedback, product reviews, and feature requests directly from comment threads.
  • 🤖 AI, NLP & LLM Training: Feed conversational datasets into retrieval-augmented generation (RAG) pipelines, sentiment classifiers, and conversational AI agents.
  • 📰 Journalism & Digital Archiving: Preserve immutable, timestamped records of notable statements, public figures' posts, and breaking news.
  • 🌟 Influencer & Creator Analytics: Track engagement velocity, follower response, and viral impact across campaigns.

Looking for industry-specific inspiration? Check out Apify's industry solutions.


📖 How to scrape Meta Threads

It's easy to scrape Meta Threads with Threads Post Scraper. Follow these simple steps:

  1. Open the Actor: Click on Try for free or open Threads Post Scraper in the Apify Console.
  2. Enter Post URLs: Paste your target Threads post URLs into the Start URLs field (accepts both threads.net and threads.com).
  3. Configure Options: The Actor is pre-configured with Apify Residential Proxies for reliable operation. Optionally set the maximum number of requests.
  4. Run the Scraper: Click Save & Start (or Run).
  5. Download Your Data: Once the run completes, preview and export your data from the Dataset tab in JSON, CSV, Excel, XML, or HTML.

📥 Input Configuration

ParameterTypeRequiredDefaultDescription
startUrlsArrayYes[{ "url": "https://www.threads.com/@openai/post/DW7RXR7EnRC" }]List of Threads post URLs to scrape (threads.com or threads.net).
proxyConfigurationObjectOptionalResidential ProxyProxy configuration for anti-bot protection. Apify Residential Proxies are recommended.
maxRequestsPerCrawlIntegerOptional100Maximum number of pages to scrape (0 = unlimited).

Example Input (input.json)

{
"startUrls": [
{ "url": "https://www.threads.com/@openai/post/DW7RXR7EnRC" },
{ "url": "https://www.threads.net/@zuck/post/C-example123" }
],
"proxyConfiguration": {
"useApifyProxy": true,
"apifyProxyGroups": ["RESIDENTIAL"]
},
"maxRequestsPerCrawl": 100
}

💰 How much will it cost to scrape Meta Threads?

Threads Post Scraper operates on a transparent, cost-effective pricing model of $1.80 per 1,000 requests ($0.0018 per request).

Apify provides $5 in free usage credits every month on the Apify Free plan. That means you can scrape up to 2,700+ Threads posts every month completely free!


📤 Results

The Actor pushes structured records to the default Apify dataset. Each item contains the primary post details in the thread object and all extracted comment items in the replies array.

Sample Output Record

{
"thread": {
"text": "There’s a new Pro tier in town. \n\nTo celebrate the launch, we’re increasing Codex usage for a limited time...",
"attachment_text": null,
"published_on": 1775770338,
"id": "3871764671238403138_63299409527",
"pk": "3871764671238403138",
"code": "DW7RXR7EnRC",
"username": "openai",
"user_pic": "https://scontent.cdninstagram.com/v/t51.82787-19/788716177_18115256264517701_n.jpg",
"user_verified": true,
"user_pk": "63299409527",
"user_id": "63299409527",
"has_audio": null,
"reply_count": 28,
"like_count": 217,
"images": ["https://instagram.fna.fbcdn.net/v/t51.82787-15/661617063_18096549797517701_n.jpg"],
"image_count": 1,
"videos": [],
"url": "https://www.threads.net/@openai/post/DW7RXR7EnRC"
},
"replies": [
{
"text": "MIGHT have to test the new tier. Looks impressive!",
"attachment_text": null,
"published_on": 1775770627,
"id": "3871767105461481006_76614717151",
"pk": "3871767105461481006",
"code": "DW7R6s-Eu4u",
"username": "tech_enthusiast",
"user_pic": "https://scontent.cdninstagram.com/v/t51.82787-19/790319655_17943256095321890_n.jpg",
"user_verified": true,
"user_pk": "76614717151",
"user_id": "76614717151",
"has_audio": null,
"reply_count": 0,
"like_count": 5,
"images": [],
"image_count": 0,
"videos": [],
"url": "https://www.threads.net/@tech_enthusiast/post/DW7R6s-Eu4u"
}
]
}

Dataset Fields Breakdown

Field NameTypeDescription
thread.textStringCaption or body text of the main post.
thread.attachment_textString | nullContent of collapsed text attachments (null if none).
thread.published_onIntegerPost publication timestamp in Unix epoch seconds.
thread.idStringFull unique Threads post identifier.
thread.codeStringAlphanumeric post shortcode (e.g., DW7RXR7EnRC).
thread.usernameStringAuthor's Threads username.
thread.user_picStringDirect URL to author's avatar image.
thread.user_verifiedBooleantrue if author account has a verified badge.
thread.like_countIntegerTotal like count registered on the post.
thread.reply_countIntegerTotal replies count registered on the post.
thread.imagesArray<String>List of direct image URLs attached to the post.
thread.videosArray<String>List of direct video URLs attached to the post.
thread.urlStringCanonical URL of the post.
repliesArray<Object>List of extracted discussion replies with matching metadata fields.

💻 Programmatic Usage & API

You can trigger Threads Post Scraper programmatically in your applications:

Python (apify-client)

from apify_client import ApifyClient
client = ApifyClient("YOUR_APIFY_TOKEN")
run_input = {
"startUrls": [{"url": "https://www.threads.com/@openai/post/DW7RXR7EnRC"}],
"maxRequestsPerCrawl": 50,
}
run = client.actor("YOUR_USERNAME/threads-post-scraper").call(run_input=run_input)
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item["thread"]["username"], ":", item["thread"]["text"])

Node.js / JavaScript (apify-client)

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: 'YOUR_APIFY_TOKEN' });
const input = {
startUrls: [{ url: 'https://www.threads.com/@openai/post/DW7RXR7EnRC' }],
maxRequestsPerCrawl: 50,
};
const run = await client.actor('YOUR_USERNAME/threads-post-scraper').call(input);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);

💡 Tips for scraping Meta Threads

  • 🎯 Use Direct Post URLs: Supply the exact post URL (https://www.threads.net/@user/post/CODE or https://www.threads.com/@user/post/CODE) for targeted and fast scraping.
  • 🛡️ Keep Residential Proxies Enabled: Threads enforces strict anti-bot protections; residential proxies ensure reliable operation and avoid IP blocks.
  • ⚙️ Set Appropriate Limits: Configure maxRequestsPerCrawl to manage data volume and stay comfortably within your monthly plan.
  • 🔄 Schedule Regular Runs: Threads data updates frequently. Schedule runs hourly or daily using Apify's built-in scheduler to track reply velocity and conversation momentum.
  • 📊 Automate Downstream Pipelines: Connect this Actor to Zapier, Make, Google Sheets, Slack, or webhooks to automatically ingest new discussions into your workflows.

Our scrapers are ethical and do not extract private user data. Threads Post Scraper extracts only publicly accessible data that Threads displays to logged-out visitors.

Please note that personal data is protected by GDPR in the European Union and by other regulations around the world. You should not scrape personal data unless you have a legitimate reason to do so. If you're unsure whether your reason is legitimate, consult your lawyers.

We also recommend that you read our comprehensive guide: Is web scraping legal?


❓ Frequently Asked Questions (FAQ)

Do I need a Meta Threads account or login credentials?

No. Threads Post Scraper works entirely without accounts, passwords, session cookies, or official API keys. It extracts only what is publicly visible to logged-out web visitors.

Does it work with both threads.com and threads.net URLs?

Yes. The scraper seamlessly handles canonical threads.net URLs as well as threads.com desktop and mobile links.

How many replies does it extract per post?

The Actor extracts all replies and nested conversations that Threads displays on the public post page. On viral posts with thousands of interactions, Threads displays the primary initial batch of replies.

Can I run this Actor via API?

Yes. Every Apify Actor can be triggered programmatically using Apify's REST API, Python SDK (apify-client), or JavaScript/TypeScript SDK.


📬 Support & Feedback

  • Bug Reports & Issues: If you notice an issue with a specific Threads post, open a ticket on the Issues tab with the affected URL.
  • Email Support: xyz.mikolabs@gmail.com
  • Rate this Actor: If you find this scraper helpful, please leave a ⭐ rating on the Apify Store!