WordPress Scraper | Posts, Content, Authors & Categories avatar

WordPress Scraper | Posts, Content, Authors & Categories

Pricing

from $0.50 / 1,000 posts

Go to Apify Store
WordPress Scraper | Posts, Content, Authors & Categories

WordPress Scraper | Posts, Content, Authors & Categories

Scrape posts or pages from any WordPress site via the public WP REST API: title, content, author, categories, tags, dates, links. No API key.

Pricing

from $0.50 / 1,000 posts

Rating

0.0

(0)

Developer

Muzaffer Kadir YILMAZ

Muzaffer Kadir YILMAZ

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

2 days ago

Last modified

Share

WordPress Scraper

Scrape posts or pages from any WordPress site by URL — title, content, excerpt, author, categories, tags, dates, featured image, and link. No API key, nothing to configure.

Paste a blog like wordpress.org/news, get its full post archive — ready to export to CSV, Excel, or JSON for content audits, migrations, or research.

Features

  • Works on most WordPress sites — just paste the site URL.
  • Posts or pages — scrape the blog archive or the site's static pages.
  • Keyword search — keep only posts matching a search term.
  • Full content — clean title, excerpt, full HTML content, author name, category and tag names.

Use cases

  • Content audits & SEO — export a site's full archive with dates and categories.
  • Competitor monitoring — track what competitors publish.
  • Migrations & backups — pull posts into another CMS or a spreadsheet.
  • Research & datasets — build text corpora from blogs and news sites.

Input

FieldTypeDescription
siteUrlsURL[]WordPress site domains or URLs.
postTypeenumposts (default) or pages.
searchstringOptional keyword filter.
maxItemsintegerMax items per site. Default 100.

Sites that have disabled public access to their content return no results.

Examples

{
"siteUrls": [{ "url": "https://wordpress.org/news" }],
"postType": "posts",
"maxItems": 100
}
{
"siteUrls": [{ "url": "https://wordpress.org/news" }],
"search": "release",
"maxItems": 50
}

Output

Each record is one post or page:

FieldDescription
site / id / typeSite host, item ID and post or page.
title / excerpt / contentTitle, short excerpt and full content.
date / modifiedPublished and last-modified dates.
author / authorIdAuthor name and ID.
categories / tagsCategory and tag names.
featuredImageFeatured image URL.
slug / status / linkURL slug, status and public link.

Example

{
"site": "wordpress.org",
"id": 21731,
"type": "posts",
"date": "2026-09-22T14:01:20",
"modified": "2026-09-23T14:04:10",
"slug": "wordpress-7-1-2-release",
"status": "publish",
"title": "WordPress 7.1.2 Release",
"excerpt": "This security release features a fix for a critical severity security vulnerability. Because this is a security release, it is recommended that you up...",
"content": "This security release features a fix for a critical severity security vulnerability. Because this is a security release, it is recommended that you up...",
"link": "https://wordpress.org/news/2026/09/wordpress-7-1-2-release/",
"author": "John Blackbourn",
"authorId": 42547,
"categories": [
"Releases",
"Security"
],
"tags": [],
"featuredImage": null,
"scrapedAt": "2026-09-23T22:03:06.019Z"
}

Download the dataset as JSON, CSV, Excel, or XML, or pull it via the Apify API.

Pricing

Pay-per-event: $0.01 per run start (per GB of run memory) plus $0.50 per 1,000 posts. You only pay for posts actually returned. See the Pricing tab for current rates.

Disclaimer

This is an unofficial Actor. It is not affiliated with or endorsed by WordPress, Automattic, or any site it reads. It returns publicly available content; you are responsible for how you use the data.