Wikipedia Page Summaries Scraper avatar

Wikipedia Page Summaries Scraper

Pricing

from $8.00 / 1,000 result items

Go to Apify Store
Wikipedia Page Summaries Scraper

Wikipedia Page Summaries Scraper

Pull Wikipedia article summaries via REST API. Returns title, description, extract (plain + HTML), thumbnail, lang, page ID, content URLs (desktop + mobile + edit), coordinates, page type, timestamps. Look up specific titles or get search results.

Pricing

from $8.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

6 days ago

Last modified

Share

ParseForge Banner

πŸ“š Wikipedia Article Summary Scraper

πŸš€ Pull Wikipedia article summaries with thumbnail, extract, coordinates, Wikidata link, revision ID, and language. Lookup or search modes.

The Wikipedia Article Summary Scraper pulls structured summaries from Wikipedia's REST API. Output includes thumbnail and original-image URLs (with widths and heights), page ID, title and display title, normalized + canonical title, description and description source, summary extract (plain text + HTML), page type, namespace, Wikibase item ID, language code and direction, last-modified timestamp, revision ID, geographic coordinates, and desktop / mobile / edit / revisions URLs.

Two modes in one Actor: lookup by title (one per line), and search (using Wikipedia's opensearch). The dataset covers Wikipedia in any of 300+ languages. Set the language input to es, fr, de, etc.

🎯 Target AudienceπŸ’‘ Primary Use Cases
Knowledge-graph builders, content marketers, ML researchers, journalists, encyclopedia apps, education platformsKnowledge-graph extraction, encyclopedic-content displays, summary embeddings, fact-card UIs, education content

πŸ“‹ What the Wikipedia Article Summary Scraper does

Five filtering workflows in a single run:

  • πŸ” Lookup mode. One title per line, returns rich summary per page.
  • πŸ” Search mode. Wikipedia's opensearch with ranked matches.
  • 🌐 300+ languages. Switch language with a single input.
  • πŸ—ΊοΈ Coordinates included. Lat / lng for places, when the page is geo-tagged.
  • πŸ”— Wikidata link. Direct Wikibase item ID for cross-language joins.

πŸ’‘ Why it matters: clean, server-side filtering and fresh data on every run.

πŸ“Š Data fields

Each record includes: description, desktopUrl, extract, language, pageId, thumbnailUrl, timestamp, title. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

πŸš€ How to use

  1. πŸ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. 🌐 Open the Actor. Find the Wikipedia Article Summary Scraper on the Apify Store.
  3. 🎯 Set input. Pick filters and maxItems.
  4. πŸš€ Run it. Click Start.
  5. πŸ“₯ Download. Grab results in the Dataset tab as CSV, Excel, JSON, or XML.

⏱️ Total time from signup to dataset: 3-5 minutes. No coding required.

πŸ’‘ Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.

⚠️ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the Wikimedia Foundation, Wikipedia editors, or any cited reference work. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.

πŸ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.