Article Content Extractor & Reader Scraper
Pricing
from $8.00 / 1,000 results
Article Content Extractor & Reader Scraper
Extract article bodies, bylines, publish dates, excerpts, and hero images from public news, blog, newsroom, and press URLs.
Article Content Extractor & Reader Scraper
Pricing
from $8.00 / 1,000 results
Extract article bodies, bylines, publish dates, excerpts, and hero images from public news, blog, newsroom, and press URLs.
Public article, news, blog, newsroom, or press URLs to extract (max 300). Route broad website pages to website-content-extractor.
Content output format. Markdown is recommended for first-run proof and LLM ingestion.
Include the inline images array in the result. The heroImage field can still be returned separately when available.
Parallel article requests (1-10).
Request timeout per article in milliseconds.
Select dataset-only output or webhook handoff. Non-dry-run always writes canonical dataset rows first; webhook delivery runs only after dataset/PPE output succeeds.
URL to POST article payloads when delivery=webhook. The webhook is sent after canonical dataset rows and PPE output succeed, and is skipped on dryRun.