Crossref Scraper - Papers, Authors & Citations
Pricing
from $1.00 / 1,000 result exporteds
Crossref Scraper - Papers, Authors & Citations
Extract academic publication metadata from Crossref: DOI, title, authors, ORCIDs, affiliations, journal, publisher, citation count, subjects and funders. Search by keyword, author, journal, publisher, work type and date range. 160M+ records, no API key required.
Pricing
from $1.00 / 1,000 result exporteds
Rating
0.0
(0)
Developer
Ryan Zinburg
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Share
Crossref Scraper - Academic Papers, Authors, Citations & Journals
Extract scholarly publication metadata from Crossref, the DOI registration agency behind most academic publishing. More than 160 million records covering journal articles, book chapters, preprints, conference papers and datasets.
Public API, no key required, no proxy needed.
What you get per publication
| Field | Example |
|---|---|
doi | 10.1016/j.mlwa.2026.101002 |
title | Machine learning model for cardiac sarcomere twitch contraction dynamics |
type | journal-article, book-chapter, posted-content, proceedings-article |
journal, journalShort | Machine Learning with Applications |
publisher | Elsevier BV |
publishedDate, publishedYear | 2026-09-01 |
citationCount | how often the work has been cited |
referencesCount | size of its own bibliography |
authors, firstAuthor, authorCount | full author list |
orcids | ORCID identifiers of the authors |
affiliations | institutions, where deposited |
issn, volume, issue, page | bibliographic details |
subjects | subject categories |
funders | funding organisations |
licenseUrl | licence of the work |
url | resolvable DOI link |
Search filters
- searchTerm - bibliographic search across title, journal and abstract, e.g.
machine learning - author - author name, e.g.
Hinton - journal - journal or container title, e.g.
Nature Communications - publisher - e.g.
Elsevier,Springer - workType - journal article, book chapter, preprint, conference paper, dataset and more
- publishedAfter / publishedBefore - publication date range
- openAccessOnly - only works with a deposited licence
- hasOrcid - only works with at least one ORCID-identified author
- sortBy - newest first, most cited first or most relevant
- maxResults - up to 10 000 works per run
Example input
{"searchTerm": "large language models","publishedAfter": "2026-01-01","sortBy": "citations","maxResults": 500}
Use cases
- Competitive research intelligence - see who publishes in your field, where and how often
- Academic recruiting - find prolific authors by topic, institution and ORCID
- Publisher and journal analytics - measure output and citation impact per journal or publisher
- Literature reviews - export a structured corpus instead of copying references by hand
- R&D scouting - track which funders and institutions back a research area
- Bibliometrics and reporting - build citation and output dashboards
Why Crossref
Crossref is the infrastructure publishers use to mint DOIs, so its metadata is the source that indexing services build on. It is open, complete for participating publishers and free to query at scale, with deep pagination through cursors rather than a hard result cap.
Notes
- Affiliations and ORCIDs are only present when the publisher deposited them; coverage is good for recent works and thinner before roughly 2015.
citationCountcounts citations recorded within Crossref, which is usually lower than Google Scholar.- Deep result sets are paged with Crossref cursors, so large exports stay reliable.
- "Newest first" orders by the date the record entered Crossref. Sorting by the printed publication date is not possible together with deep paging, and that field contains future-dated issues.