Project Gutenberg Books Scraper | 70K+ Free eBooks
Pricing
from $19.00 / 1,000 result items
Project Gutenberg Books Scraper | 70K+ Free eBooks
Export 70,000+ public-domain books from Project Gutenberg via the Gutendex API. Search by keyword, language, topic, or author lifespan, or fetch by book ID. Pull titles, authors, subjects, languages, download links, and full-text formats. Download as CSV, Excel, JSON, or XML.
Pricing
from $19.00 / 1,000 result items
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Categories
Share

๐ Project Gutenberg (Gutendex) Scraper
๐ Export 70,000+ public-domain books with metadata and full-text download links in seconds.
This Apify Actor extracts structured data from Project Gutenberg (Gutendex), returning clean JSON / CSV / Excel / XML datasets ready for analytics, integrations, or research workflows. Built by ParseForge for reliability and freshness.
| ๐ฏ Target Audience | ๐ก Primary Use Cases |
|---|---|
| Data analysts, engineers, researchers | Analytics pipelines, BI dashboards, datasets |
| SaaS, fintech, marketing, ops teams | Lead gen, enrichment, monitoring |
| Hobbyists, journalists, indie devs | Side projects, content, exploration |
๐ What the Project Gutenberg (Gutendex) Scraper does
- Queries the public Project Gutenberg (Gutendex) API / feed and structures the response
- Returns one record per item with 10 normalized fields
- Supports filters configurable from the input schema
- Outputs to CSV, Excel, JSON, XML via Apify dataset
- Auto-limits to 10 items on the free plan; up to 1,000,000 on paid
๐ก Why it matters: clean, ready-to-query data without manual scraping, parsing, or babysitting an API client.
๐ Data fields
Each record includes: allFormats, authors, bookId, bookshelves, copyright, coverImageUrl, downloadCount, editors, epubUrl, htmlUrl, kindleUrl, languages, mediaType, plainTextUrl, readOnlineUrl, scrapedAt, subjects, summary, title, translators, url. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
๐ How to use
- Create a free Apify account with $5 credit
- Open this Actor and click Try for free
- Configure the input (
maxItemsand any filters) - Click Start
- Download the dataset as CSV / Excel / JSON / XML
๐ Recommended Actors
| Actor | Description |
|---|---|
| Wikipedia On This Day Scraper | Daily Wikipedia historical events |
| Public Holidays Scraper | Holidays for 100+ countries |
| SPDX Software Licenses Scraper | Open-source license metadata |
| ISO Country Codes Scraper | IBAN + ISO country codes |
๐ก Pro Tip: browse the complete ParseForge collection.
โ ๏ธ Disclaimer: independent tool, not affiliated with Project Gutenberg (Gutendex). Only publicly available data collected.
๐ Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.