ArXiv Paper Scraper
Pricing
from $1.00 / 1,000 arxiv paper scrapers
ArXiv Paper Scraper
Search arXiv and export structured paper metadata — title, authors, abstract, categories, DOI, PDF link — via the official arXiv API. No key, no anti-bot. Pay per paper.
Pricing
from $1.00 / 1,000 arxiv paper scrapers
Rating
0.0
(0)
Developer
Pedro Resende
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
3 days ago
Last modified
Categories
Share
ArXiv Paper Scraper searches arXiv and returns rich, structured paper metadata — title, authors, abstract, categories, DOI, journal reference and PDF link — using the official arXiv API (Atom XML). No API key, no anti-bot, no rental. Pay only per paper.
What this arXiv scraper does
- Search arXiv by keyword (title, author, abstract, category) and get every matching paper with full metadata.
- Fetch specific papers by arXiv ID (
2106.15928,cs/0101001). - Filter by category (
cs.AI,cs.LG,cs.CL,math.OC,quant-ph, …). - Sort by relevance, last-updated or submission date, and page with
start. - Great for literature review, citation datasets, competitor research, and feeding papers into AI pipelines.
Input
| Field | Description |
|---|---|
query | Search terms (supports arXiv syntax: all:, ti:, au:, abs:). |
idList | Fetch specific papers by arXiv ID. |
category | Restrict to one arXiv category, e.g. cs.AI. |
maxResults | Max papers to return (1–100). |
start | Pagination offset (0-based). |
sortBy | relevance, lastUpdatedDate, or submittedDate. |
sortOrder | descending or ascending. |
Output
Each dataset item includes id, title, summary (abstract), authors[], categories[], primaryCategory, doi, journalRef, comment, published, updated, absUrl and pdfUrl.
Pricing
Pay per event: billed once per paper returned. An aborted run only pays for what it delivered. No monthly rental.
Note: this Actor returns arXiv metadata via the official arXiv API — no scraping of third-party PDF hosts.