Europe PMC Literature Scraper
Pricing
from $27.60 / 1,000 results
Europe PMC Literature Scraper
Scrape Europe PMC for biomedical research papers. Search by title, author, MeSH terms, journal. Get DOI, abstract, full-text URLs, citations, references, open-access status. No API key required.
Pricing
from $27.60 / 1,000 results
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
4
Total users
0
Monthly active users
21 hours ago
Last modified
Categories
Share

๐งฌ Europe PMC Literature Scraper
๐ Export the biomedical literature index in seconds. Search 40+ million records across PubMed, PubMed Central, life-science preprints, agricultural literature, and patents. Filter by title, author, MeSH term, DOI, journal, open access, or free-text. No API key, no registration.
The Europe PMC Literature Scraper wraps the official Europe PMC REST API (ebi.ac.uk/europepmc/webservices/rest/search) and returns one row per article with 40+ fields, including DOI, PMID, PMCID, abstract, full-text URLs, MeSH terms, keywords, journal, citation count, open-access status, and licensing. The underlying corpus is published by Europe PMC, the European mirror of PubMed Central, maintained by EMBL-EBI and funded by 32 life-science research funders worldwide.
The index covers MEDLINE/PubMed, PubMed Central (full text), Agricola (USDA agricultural literature), bioRxiv and medRxiv preprints, CTX patents, and Europe PMC-curated content. Free-text and field-qualified queries (TITLE, AUTH, MESH, DOI, PMID, AFFILIATION, JOURNAL) compose freely with boolean operators. This Actor returns structured records ready to download as CSV, Excel, JSON, or XML.
| ๐ฏ Target Audience | ๐ก Primary Use Cases |
|---|---|
| Biomedical researchers, systematic-review teams, bibliometrics analysts, pharma intelligence, scientific publishers, science journalists, OA advocacy, ML training pipelines | Literature reviews, MeSH-term mining, author publication tracking, journal impact studies, drug-target evidence harvesting, training-set assembly |
๐ What the Europe PMC Scraper does
One programmable interface to the full Europe PMC search service:
- ๐ Field-qualified queries.
TITLE:,AUTH:,AFFILIATION:,JOURNAL:,MESH:,DOI:,PMID:,PMCID:,OPEN_ACCESS:, plus boolean operators (AND,OR,NOT) and quoted phrases. - ๐ Three response shapes.
corereturns the full record with abstract, full-text URLs, and metadata.litereturns compact fields.idlistreturns IDs only for ultra-fast scans. - โฑ๏ธ Sort options. Relevance (default), newest first, oldest first, or most cited.
- ๐ Cursor-mark pagination. Fully automatic. Walks the entire result set efficiently for large queries.
Output captures the publication metadata (PMID, PMCID, DOI, source, journal title, ISSN, volume, issue, page info, publication year and date), full author list, abstract text, affiliation, language, publication types, MeSH headings, keywords, grant count, citation count, full-text URLs, license, open-access flag, and indexing dates.
๐ก Why it matters: Europe PMC is the deepest open-access biomedical literature index in the world. The web UI is great for one-off lookups, but systematic reviews, bibliometric studies, and ML training-set assembly need flat rows. This Actor turns the search service into a downloadable dataset in one run.
๐ Data fields
Each record includes: abstractText, affiliation, authorList, authorString, citedByCount, dateOfRevision, doi, firstIndexDate, firstPublicationDate, fullTextUrls, grantsCount, hasBook, hasDbCrossReferences, hasPDF, hasReferences, hasSuppl, hasTextMinedTerms, id, inEPMC, inPMC, isOpenAccess, issue, journalIssn, journalTitle, journalVolume, keywords, language, license, meshTerms, pageInfo, pmcid, pmid, pubDate, pubYear, publicationStatus, publicationTypes, scrapedAt, source, title, url. All 40 field names come from a real production run, so what you see here is what lands in your dataset.
๐ How to use
- ๐ Sign up. Create a free account with $5 credit (takes 2 minutes).
- ๐ Open the Actor. Go to the Europe PMC Literature Scraper page on the Apify Store.
- ๐ Build a query. Free-text or use field qualifiers (
AUTH:"Doudna J" AND CRISPR). - ๐ Pick a response shape.
corefor full metadata,litefor compact,idlistfor IDs only. - ๐ Run it. Click Start and let the Actor collect your data.
- ๐ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.
โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.
๐ Recommended Actors
- ๐ค Hugging Face Model Scraper - ML model registry metadata
- ๐ช๐บ Eurostat Statistics Scraper - 7,500+ Eurostat datasets
- ๐ ClinicalTrials.gov Scraper - Clinical trial registry
- ๐ Figshare Research Output Scraper - Open research datasets
- ๐ฌ OSF Open Science Framework Scraper - Open-science project metadata
๐ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.
โ ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Europe PMC, EMBL-EBI, the European Bioinformatics Institute, the National Center for Biotechnology Information, or any of the 32 funders supporting Europe PMC. All trademarks mentioned are the property of their respective owners. Only publicly available open data from the official Europe PMC REST API is collected.
๐ Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.