arXiv Research Papers & Abstracts Scraper avatar

arXiv Research Papers & Abstracts Scraper

Pricing

from $2.55 / 1,000 results

Go to Apify Store
arXiv Research Papers & Abstracts Scraper

arXiv Research Papers & Abstracts Scraper

Scrape arXiv preprints by keyword, author or subject with arXiv ID, title, authors, abstract, subject categories, DOI, publication and update dates and PDF links. Export to JSON, CSV or Excel.

Pricing

from $2.55 / 1,000 results

Rating

0.0

(0)

Developer

Scrapers Lat

Scrapers Lat

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

4 days ago

Last modified

Share

Search arXiv and export clean, structured paper metadata: arXiv ID, title, authors, full abstract, subject categories, DOI, publication and update dates and PDF links. Perfect for literature reviews, research dashboards and citation tracking.

📥 Input · 📤 Output · 💰 Pricing · ▶️ Examples

Apify Research papers Output

Full abstracts
plus authors
Categories & DOI
with PDF links
JSON / CSV / Excel
output formats

What you get

Each record is one arXiv paper, ready for a reference manager, a dashboard or a spreadsheet:

  • arxivId: the arXiv identifier (with version)
  • title: the paper title
  • authors and authorCount: the full author list
  • summary: the complete abstract
  • categories and primaryCategory: arXiv subject classifications
  • published and updated: submission and last revision dates
  • doi: the DOI when the paper has one
  • journalRef: journal reference when available
  • comment: author-supplied notes (pages, figures, conference)
  • pdfUrl and absUrl: direct links to the PDF and the abstract page
  • observedAt: when the record was collected

Who is it for

Use caseWho benefits
Literature reviewsResearchers building a corpus on a topic fast
Research dashboardsTeams tracking new work in a field
Citation and trend analysisAnalysts studying authors, categories and volume over time
ML datasetsBuilders assembling training or evaluation sets of abstracts

How to use it

  1. Enter a search query (for example large language models, quantum computing or protein folding).
  2. Optionally choose which field to match (title, abstract, author or category) and how to sort (relevance, newest, recently updated).
  3. Set Max Items and run. Export as JSON, CSV or Excel, or pull it through the Apify API.

Frequently Asked Questions

Can I search by author or subject category? Yes. Set the search field to Author or Category code, or pass a native query such as au:hinton or cat:cs.CL. You can also combine terms with AND / OR.

Do I get the full abstract? Yes. Each record includes the complete abstract text, not just a snippet.

Does every paper have a DOI? No. A DOI is included whenever the paper has one registered; otherwise the field is null. Every record still has a stable arXiv ID and links.

How many papers can I collect? Set Max Items to whatever you need. Results are gathered page by page until that limit or the end of the matches is reached.

Example use cases

Ready-to-run example tasks, each preconfigured for a common scenario. Open one and press run, or use it as a template:

More scrapers at scrapers.lat

This actor is built and maintained by scrapers.lat, where we publish scrapers for public platforms: finance, news, real estate, jobs, e-commerce and government data. Browse the full catalog or ask us for a custom scraper at scrapers.lat.


This actor is an independent tool and has no affiliation with arXiv or Cornell University. It only accesses publicly available paper metadata. Use the results in accordance with the source's terms.