Open Library Scraper avatar

Open Library Scraper

Pricing

Pay per event

Go to Apify Store
Open Library Scraper

Open Library Scraper

Comprehensive scraper for Open Library to extract books, authors, subjects, and list data from the Internet Archive’s platform. Supports multiple search types and ebook filtering, providing automated, structured access to Open Library’s extensive bibliographic collection.

Pricing

Pay per event

Rating

5.0

(1)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

1

Bookmarked

15

Total users

2

Monthly active users

16 hours ago

Last modified

Share

ParseForge Banner

πŸ“š Open Library Scraper

πŸš€ Extract book data from Open Library in seconds. Search by title, author, or subject with ebook filtering. No coding, no API keys required.

Open Library is the Internet Archive's free, open catalog of every book ever published. This scraper connects to Open Library's public API and returns structured book data including titles, authors, ISBNs, publishers, cover images, ratings, descriptions, page counts, and download links. It supports 5 search types (Books, Authors, Search Inside, Subjects, and Lists), handles pagination automatically, and exports data as JSON, CSV, or Excel.

Whether you are building a book recommendation engine, tracking ISBN availability, or conducting bibliographic research, this actor delivers structured data for up to 1,000,000 records per run for paid users. Each result includes cover images, publication dates, publisher names, available formats (PDF, EPUB, AZW3), community ratings, edition counts, and subject classifications. No manual searching, copying, or format conversion needed.

🎯 Target AudienceπŸ’‘ Use Cases
LibrariansBuild digital catalogs with cover images and ISBNs
PublishersResearch publication history and edition counts
Book bloggersGenerate reading lists with ratings and descriptions
Data scientistsAnalyze publishing trends by subject and year
App developersFeed book metadata into recommendation engines
EducatorsCurate subject-specific reading lists for courses

πŸ“‹ What the Open Library Scraper does

  • πŸ” Keyword search across books, authors, subjects, lists, and full-text content
  • πŸ“– Ebook filtering to show only books available as free digital downloads
  • πŸ–ΌοΈ Cover image extraction with URLs for small, medium, and large sizes
  • πŸ“Š Edition and rating data including community ratings and total edition counts
  • πŸ“₯ Download link collection for PDF, EPUB, and AZW3 formats when available
  • 🌐 Direct URL support to scrape any Open Library search results page

The scraper sends your query to Open Library's public API, retrieves matching records, and extracts full metadata for each item. For book searches, it collects titles, authors, ISBNs, publishers, page counts, descriptions, cover images, ratings, available formats, and download links. For author searches, it returns author profiles with their works. Every record is timestamped and includes a direct link to the Open Library entry.

πŸ’‘ Why it matters: Open Library contains metadata for millions of books, but browsing and exporting data manually is tedious. This scraper automates collection and delivers clean, structured data ready for databases, spreadsheets, or applications.

πŸ“Š Data fields

Each record includes: author, availableFormats, coverImages, detailUrl, downloadLinks, editionCount, fullDescription, imageUrl, isbn, itemId, language, numberOfPages, publicationDate, publishers, rating, scrapedTimestamp, subjectTags, subjects, title. All 19 field names come from a real production run, so what you see here is what lands in your dataset.

⚠️ Good to Know: Use either a Start URL or search filters, not both. If you provide a Start URL, search filters are ignored. The ebooks-only filter only works with the "books" search type.

πŸš€ How to use

  1. Create an Apify account - Sign up free with $5 credit
  2. Open the Open Library Scraper - Navigate to the actor page on Apify
  3. Enter your search query - Type a title, author name, or subject
  4. Select search type and filters - Choose Books, Authors, Subjects, etc. and enable ebook filtering if needed
  5. Click Start - The actor collects matching records and delivers structured data

⏱️ A typical run with 10 books completes in under 30 seconds.

ActorDescription
PubMed Citation ScraperExtract publication metadata from PubMed for research analysis
Crossref ScraperExtract DOI metadata for 155M+ research publications
NASA Reports ScraperCollect technical reports from NASA's NTRS database
US Census Bureau ScraperExtract demographic and economic data from the Census Bureau
ROR ScraperCollect research organization data from the Research Organization Registry

πŸ’‘ Pro Tip: Combine the Open Library Scraper with the Crossref Scraper to match book ISBNs with DOI metadata and citation counts.

Disclaimer: This actor is not affiliated with, endorsed by, or connected to Open Library or the Internet Archive. It accesses publicly available data through Open Library's public API. Use responsibly and in accordance with applicable terms of service.

πŸ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.