Wikidata Entity Search Scraper avatar

Wikidata Entity Search Scraper

Pricing

from $14.00 / 1,000 result items

Go to Apify Store
Wikidata Entity Search Scraper

Wikidata Entity Search Scraper

Search Wikidata's open knowledge graph of 100M+ entities (people, places, brands, books, films) by name. Returns Q-ID, label, description, aliases, all claims (P-properties), sitelinks to every Wikipedia language, structured facts and image. Filter by entity type, language and full-claims fetching.

Pricing

from $14.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

4

Total users

1

Monthly active users

3 days ago

Last modified

Share

ParseForge Banner

๐ŸŒ Wikidata Entity Search Scraper

๐Ÿš€ Search Wikidata's open knowledge graph of 100M+ entities by name.

The Wikidata Entity Search Scraper searches Wikidata's open knowledge graph of 100M+ entities by name. Output includes the canonical Q-ID, label, description, aliases, all claims (P-properties), sitelinks to every Wikipedia language edition, and structured facts.

Wikidata is the structured-data backbone of Wikipedia and one of the largest open knowledge graphs in the world. Filters run server-side, so a single run can resolve every entity matching a name, fetch full claim trees, or pull entities in non-English languages.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
ML pipelines, knowledge-graph engineers, journalists, fact-checkers, content recommendation engines, search developersEntity resolution, knowledge-graph augmentation, fact-checking, content enrichment, multilingual search, ML training datasets

๐Ÿ“‹ What the Wikidata Entity Search Scraper does

Five filtering workflows in a single run:

  • ๐Ÿ” Free-text search. Match entity labels and aliases.
  • ๐ŸŒ Multilingual. Search in 20+ languages (en, es, fr, de, it, ja, zh, ko, ar, hi, pt, nl, ru).
  • ๐Ÿ†” Item or property. Search Q-entities (items) or P-entities (properties).
  • ๐Ÿ“Š Full claims fetch. Optional: pull every statement, sitelink, and structured fact per entity.
  • ๐Ÿท๏ธ Image extraction. Auto-extracts the entity's primary image from claim P18.

๐Ÿ’ก Why it matters: clean, server-side filtering removes the parser-and-pagination work from your team and keeps your dataset fresh on every run.

๐Ÿ“Š Data fields

Each record includes: aliasesText, claimCount, description, entityId, instanceOf, label, sitelinkCount, thumbnailUrl, wikidataUrl, wikipediaEnUrl. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Go to the Wikidata Entity Search Scraper page on the Apify Store.
  3. ๐ŸŽฏ Set input. Pick your filters and maxItems.
  4. ๐Ÿš€ Run it. Click Start and let the Actor collect your data.
  5. ๐Ÿ“ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.

โš ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the Wikimedia Foundation, Wikidata, Wikipedia, or any contributing editor. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.