Go to example tasks
Full Wikipedia Article Text for RAG & Knowledge Bases
The complete plain text of Wikipedia articles (here: Machine learning, Large language model, Artificial intelligence) with headings kept, a table of contents, description, Wikidata id and URL. For AI engineers building RAG pipelines, chatbots and knowledge bases. Replace the queries, or set lang (de, fr, es, ja…) to read another Wikipedia edition.
Wikipedia Scraper — Search, Summaries & Full Textyadroo/wikipedia-search
Query
Title
Description
Extract
+4 fieldsTextNumberBooleanListObject
Input
Search queries or page titles(required):Machine learning+2
Language edition:en
Text content:full
Output fields
Query
Title
Description
Extract
Lang
Wikidata
Last edited
URL
Sign up on Apify01
Create your Apify account to access the Wikipedia Scraper — Search, Summaries & Full Text.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
