ArXiv Paper Scraper avatar

ArXiv Paper Scraper

Pricing

from $1.00 / 1,000 arxiv paper scrapers

Go to Apify Store
ArXiv Paper Scraper

ArXiv Paper Scraper

Search arXiv and export structured paper metadata — title, authors, abstract, categories, DOI, PDF link — via the official arXiv API. No key, no anti-bot. Pay per paper.

Pricing

from $1.00 / 1,000 arxiv paper scrapers

Rating

0.0

(0)

Developer

Pedro Resende

Pedro Resende

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Share

ArXiv Paper Scraper searches arXiv and returns rich, structured paper metadata — title, authors, abstract, categories, DOI, journal reference and PDF link — using the official arXiv API (Atom XML). No API key, no anti-bot, no rental. Pay only per paper.

What this arXiv scraper does

  • Search arXiv by keyword (title, author, abstract, category) and get every matching paper with full metadata.
  • Fetch specific papers by arXiv ID (2106.15928, cs/0101001).
  • Filter by category (cs.AI, cs.LG, cs.CL, math.OC, quant-ph, …).
  • Sort by relevance, last-updated or submission date, and page with start.
  • Great for literature review, citation datasets, competitor research, and feeding papers into AI pipelines.

Input

FieldDescription
querySearch terms (supports arXiv syntax: all:, ti:, au:, abs:).
idListFetch specific papers by arXiv ID.
categoryRestrict to one arXiv category, e.g. cs.AI.
maxResultsMax papers to return (1–100).
startPagination offset (0-based).
sortByrelevance, lastUpdatedDate, or submittedDate.
sortOrderdescending or ascending.

Output

Each dataset item includes id, title, summary (abstract), authors[], categories[], primaryCategory, doi, journalRef, comment, published, updated, absUrl and pdfUrl.

Pricing

Pay per event: billed once per paper returned. An aborted run only pays for what it delivered. No monthly rental.

Note: this Actor returns arXiv metadata via the official arXiv API — no scraping of third-party PDF hosts.