DOAB Directory of Open Access Books Scraper avatar

DOAB Directory of Open Access Books Scraper

Pricing

from $11.00 / 1,000 result items

Go to Apify Store
DOAB Directory of Open Access Books Scraper

DOAB Directory of Open Access Books Scraper

Browse the Directory of Open Access Books (DOAB) with peer-reviewed academic titles. Capture title, authors, publisher, ISBN, DOI, subjects, language, publication year, abstract, and download URL. Export to JSON, CSV, or Excel for libraries, researchers, and content aggregation.

Pricing

from $11.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

6 days ago

Last modified

Share

ParseForge Banner

๐Ÿ“š DOAB Directory of Open Access Books Scraper

๐Ÿš€ Export the global open-access book catalog in seconds. Pull 70,000+ peer-reviewed academic titles across 25+ subject areas in 50+ languages. No API key, no registration, no manual catalog scraping.

The DOAB Scraper exports the Directory of Open Access Books, a community-maintained catalog of peer-reviewed scholarly monographs and edited volumes that anyone can read, download, and redistribute. Each record carries 16 fields with authors, publishers, subjects, licenses, ISBNs, DOIs, abstracts, and direct download links to the full text. The underlying catalog is curated by libraries and publishers worldwide and is one of the most cited open-access references in higher education.

Coverage spans the humanities, social sciences, STEM, law, and the arts across 70,000+ titles, 700+ publishers, and 50+ languages. Every book is released under a Creative Commons or equivalent open license. This Actor turns that catalog into a CSV, Excel, JSON, or XML download in under five minutes.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
Academic librarians, OER advocates, researchers, university publishers, digital humanities labs, repository managersLibrary catalog enrichment, OER course design, bibliometric studies, repository ingest, open-access discovery, syllabus building

๐Ÿ“‹ What the DOAB Scraper does

Four discovery workflows in a single run:

  • ๐Ÿ” Full-text search. Query titles, authors, and abstracts across the entire catalog.
  • ๐Ÿ—ฃ๏ธ Language filter. Restrict to English, German, Spanish, French, or any of 50+ languages.
  • ๐Ÿ›๏ธ Subject filter. Narrow to philosophy, history, mathematics, sociology, and 25+ disciplines.
  • ๐Ÿ†” Direct handle lookup. Pull a single book by its DOAB handle when you already know the ID.

Each record includes UUID, DOAB handle, title, authors, publisher, language, subjects, ISBNs, DOI, abstract, license, publication date, and direct PDF/EPUB download URLs.

๐Ÿ’ก Why it matters: open-access monographs are the backbone of modern OER programs and digital library collections. Building your own ingest pipeline means dealing with OAI-PMH, MARC, Dublin Core mappings, and inconsistent metadata. This Actor returns a clean, normalized record on every run.

๐Ÿ“Š Data fields

Each record includes: abstract, authors, detailUrl, doi, downloadUrls, handle, isbn, language, lastModified, license, publicationDate, publisher, scrapedAt, subjects, title, uuid. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Go to the DOAB Directory of Open Access Books Scraper page on the Apify Store.
  3. ๐ŸŽฏ Set input. Pick a subject, language, or publisher (or leave defaults for a wide pull) and set maxItems.
  4. ๐Ÿš€ Run it. Click Start and let the Actor collect your data.
  5. ๐Ÿ“ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more open-data and research scrapers.

โš ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by DOAB, OAPEN, or any of its contributing publishers. All trademarks mentioned are the property of their respective owners. Only publicly available open-access catalog data is collected.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.