Open Library Scraper
Pricing
Pay per event
Open Library Scraper
Comprehensive scraper for Open Library to extract books, authors, subjects, and list data from the Internet Archiveβs platform. Supports multiple search types and ebook filtering, providing automated, structured access to Open Libraryβs extensive bibliographic collection.
Pricing
Pay per event
Rating
5.0
(1)
Developer
ParseForge
Maintained by CommunityActor stats
1
Bookmarked
15
Total users
2
Monthly active users
16 hours ago
Last modified
Categories
Share

π Open Library Scraper
π Extract book data from Open Library in seconds. Search by title, author, or subject with ebook filtering. No coding, no API keys required.
Open Library is the Internet Archive's free, open catalog of every book ever published. This scraper connects to Open Library's public API and returns structured book data including titles, authors, ISBNs, publishers, cover images, ratings, descriptions, page counts, and download links. It supports 5 search types (Books, Authors, Search Inside, Subjects, and Lists), handles pagination automatically, and exports data as JSON, CSV, or Excel.
Whether you are building a book recommendation engine, tracking ISBN availability, or conducting bibliographic research, this actor delivers structured data for up to 1,000,000 records per run for paid users. Each result includes cover images, publication dates, publisher names, available formats (PDF, EPUB, AZW3), community ratings, edition counts, and subject classifications. No manual searching, copying, or format conversion needed.
| π― Target Audience | π‘ Use Cases |
|---|---|
| Librarians | Build digital catalogs with cover images and ISBNs |
| Publishers | Research publication history and edition counts |
| Book bloggers | Generate reading lists with ratings and descriptions |
| Data scientists | Analyze publishing trends by subject and year |
| App developers | Feed book metadata into recommendation engines |
| Educators | Curate subject-specific reading lists for courses |
π What the Open Library Scraper does
- π Keyword search across books, authors, subjects, lists, and full-text content
- π Ebook filtering to show only books available as free digital downloads
- πΌοΈ Cover image extraction with URLs for small, medium, and large sizes
- π Edition and rating data including community ratings and total edition counts
- π₯ Download link collection for PDF, EPUB, and AZW3 formats when available
- π Direct URL support to scrape any Open Library search results page
The scraper sends your query to Open Library's public API, retrieves matching records, and extracts full metadata for each item. For book searches, it collects titles, authors, ISBNs, publishers, page counts, descriptions, cover images, ratings, available formats, and download links. For author searches, it returns author profiles with their works. Every record is timestamped and includes a direct link to the Open Library entry.
π‘ Why it matters: Open Library contains metadata for millions of books, but browsing and exporting data manually is tedious. This scraper automates collection and delivers clean, structured data ready for databases, spreadsheets, or applications.
π Data fields
Each record includes: author, availableFormats, coverImages, detailUrl, downloadLinks, editionCount, fullDescription, imageUrl, isbn, itemId, language, numberOfPages, publicationDate, publishers, rating, scrapedTimestamp, subjectTags, subjects, title. All 19 field names come from a real production run, so what you see here is what lands in your dataset.
β οΈ Good to Know: Use either a Start URL or search filters, not both. If you provide a Start URL, search filters are ignored. The ebooks-only filter only works with the "books" search type.
π How to use
- Create an Apify account - Sign up free with $5 credit
- Open the Open Library Scraper - Navigate to the actor page on Apify
- Enter your search query - Type a title, author name, or subject
- Select search type and filters - Choose Books, Authors, Subjects, etc. and enable ebook filtering if needed
- Click Start - The actor collects matching records and delivers structured data
β±οΈ A typical run with 10 books completes in under 30 seconds.
π Recommended Actors
| Actor | Description |
|---|---|
| PubMed Citation Scraper | Extract publication metadata from PubMed for research analysis |
| Crossref Scraper | Extract DOI metadata for 155M+ research publications |
| NASA Reports Scraper | Collect technical reports from NASA's NTRS database |
| US Census Bureau Scraper | Extract demographic and economic data from the Census Bureau |
| ROR Scraper | Collect research organization data from the Research Organization Registry |
π‘ Pro Tip: Combine the Open Library Scraper with the Crossref Scraper to match book ISBNs with DOI metadata and citation counts.
Disclaimer: This actor is not affiliated with, endorsed by, or connected to Open Library or the Internet Archive. It accesses publicly available data through Open Library's public API. Use responsibly and in accordance with applicable terms of service.
π Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.