Wikipedia Page Summaries Scraper
Pricing
from $8.00 / 1,000 result items
Wikipedia Page Summaries Scraper
Pull Wikipedia article summaries via REST API. Returns title, description, extract (plain + HTML), thumbnail, lang, page ID, content URLs (desktop + mobile + edit), coordinates, page type, timestamps. Look up specific titles or get search results.
Pricing
from $8.00 / 1,000 result items
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
6 days ago
Last modified
Share

π Wikipedia Article Summary Scraper
π Pull Wikipedia article summaries with thumbnail, extract, coordinates, Wikidata link, revision ID, and language. Lookup or search modes.
The Wikipedia Article Summary Scraper pulls structured summaries from Wikipedia's REST API. Output includes thumbnail and original-image URLs (with widths and heights), page ID, title and display title, normalized + canonical title, description and description source, summary extract (plain text + HTML), page type, namespace, Wikibase item ID, language code and direction, last-modified timestamp, revision ID, geographic coordinates, and desktop / mobile / edit / revisions URLs.
Two modes in one Actor: lookup by title (one per line), and search (using Wikipedia's opensearch). The dataset covers Wikipedia in any of 300+ languages. Set the language input to es, fr, de, etc.
| π― Target Audience | π‘ Primary Use Cases |
|---|---|
| Knowledge-graph builders, content marketers, ML researchers, journalists, encyclopedia apps, education platforms | Knowledge-graph extraction, encyclopedic-content displays, summary embeddings, fact-card UIs, education content |
π What the Wikipedia Article Summary Scraper does
Five filtering workflows in a single run:
- π Lookup mode. One title per line, returns rich summary per page.
- π Search mode. Wikipedia's opensearch with ranked matches.
- π 300+ languages. Switch language with a single input.
- πΊοΈ Coordinates included. Lat / lng for places, when the page is geo-tagged.
- π Wikidata link. Direct Wikibase item ID for cross-language joins.
π‘ Why it matters: clean, server-side filtering and fresh data on every run.
π Data fields
Each record includes: description, desktopUrl, extract, language, pageId, thumbnailUrl, timestamp, title. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
π How to use
- π Sign up. Create a free account with $5 credit (takes 2 minutes).
- π Open the Actor. Find the Wikipedia Article Summary Scraper on the Apify Store.
- π― Set input. Pick filters and
maxItems. - π Run it. Click Start.
- π₯ Download. Grab results in the Dataset tab as CSV, Excel, JSON, or XML.
β±οΈ Total time from signup to dataset: 3-5 minutes. No coding required.
π Recommended Actors
- π Wikidata Entity Search - 100M+ open knowledge-graph entities
- π Wikivoyage Travel Articles - Wikivoyage city and country articles with image, geo
- π REST Countries Reference Data - Every country with flag, capital, currency, languages
- π Stack Exchange Questions - Search 170+ Stack Exchange Q&A sites
- π GeoNames Places + Postal Codes - 12M+ places with admin hierarchy, lat/lng, alternate names
π‘ Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.
β οΈ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the Wikimedia Foundation, Wikipedia editors, or any cited reference work. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.
π Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.