Go to example tasks
Substack Pragmatic Engineer Free Posts
Crawls the Pragmatic Engineer archive, skips paywalled posts, and returns up to 20 free posts with title, subtitle, author, date, full text, likes and comment counts and URL. A cache limits repeat runs to new posts. Useful for tracking engineering industry commentary and building a software-topic reading corpus.
Substack Scraper — Posts, Comments & Paraphrase with AIconfidential_gnat/substack-newsletter-scraper
Title
Subtitle
Authors
Publication
+11 fieldsTextNumberBooleanListObject
Input
Start URLs
url:https://newsletter.pragmaticengineer.com/archive
Max posts per start URL:20
Include paywalled posts:false
Cache project name (scrape only new posts):substack-pragmatic-engineer-free-posts
Scrape comments:false
Max comments per post:5
Enable AI enrichment:false
AI features:summarize+2
AI models:openai/gpt-4o-mini
Proxy configuration
Output fields
Title
Subtitle
Authors
Publication
Published at
Audience
Is paywalled
Body
Word count
Tags
Cover image
Reaction count
Comment count
Restack count
Url
Sign up on Apify01
Create your Apify account to access the Substack Scraper — Posts, Comments & Paraphrase with AI.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
