arXiv Research Papers & Abstracts Scraper avatar

arXiv Research Papers & Abstracts Scraper

Pricing

from $6.80 / 1,000 results

Go to Apify Store
arXiv Research Papers & Abstracts Scraper

arXiv Research Papers & Abstracts Scraper

Scrape arXiv preprints by keyword, author or subject with arXiv ID, title, authors, abstract, subject categories, DOI, publication and update dates and PDF links. Export to JSON, CSV or Excel.

Pricing

from $6.80 / 1,000 results

Rating

0.0

(0)

Developer

Scrapers Lat

Scrapers Lat

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

20 hours ago

Last modified

Share

Search arXiv and export clean, structured paper metadata: arXiv ID, title, authors, full abstract, subject categories, DOI, publication and update dates and PDF links. Perfect for literature reviews, research dashboards and citation tracking.

📥 Input · 📤 Output · 💰 Pricing · ▶️ Examples

Apify Research papers Output

Full abstracts
plus authors
Categories & DOI
with PDF links
JSON / CSV / Excel
output formats

Who is it for

Use caseWho benefits
Literature reviewsResearchers building a corpus on a topic fast
Research dashboardsTeams tracking new work in a field
Citation and trend analysisAnalysts studying authors, categories and volume over time
ML datasetsBuilders assembling training or evaluation sets of abstracts

How to use it

  1. Enter a search query (for example large language models, quantum computing or protein folding).
  2. Optionally choose which field to match (title, abstract, author or category) and how to sort (relevance, newest, recently updated).
  3. Set Max Items and run. Export as JSON, CSV or Excel, or pull it through the Apify API.

Frequently Asked Questions

Can I search by author or subject category? Yes. Set the search field to Author or Category code, or pass a native query such as au:hinton or cat:cs.CL. You can also combine terms with AND / OR.

Do I get the full abstract? Yes. Each record includes the complete abstract text, not just a snippet.

Does every paper have a DOI? No. A DOI is included whenever the paper has one registered; otherwise the field is null. Every record still has a stable arXiv ID and links.

How many papers can I collect? Set Max Items to whatever you need. Results are gathered page by page until that limit or the end of the matches is reached.

Example use cases

Ready-to-run example tasks, each preconfigured for a common scenario. Open one and press run, or use it as a template:

Export, API and AI agents (x402 + MCP)

Export the scraped data to JSON, CSV or Excel, pull it as a dataset through the Apify API, or wire it into your app with no code. This web scraper and data extractor also works for bulk data extraction and scheduled runs.

For AI agents: this Actor is available on x402, Apify's agentic payment standard built with Coinbase. An AI agent can discover, pay for and run it on its own with a funded wallet and a single HTTP request: no account, no subscription, no API key and no human in the loop. It also runs as an MCP tool inside Claude, Cursor and other AI clients out of the box. Learn more about x402 agentic payments on Apify.

More scrapers at scrapers.lat

This actor is built and maintained by scrapers.lat, where we publish scrapers for public platforms: finance, news, real estate, jobs, e-commerce and government data. Browse the full catalog or ask us for a custom scraper at scrapers.lat.


This actor is an independent tool and has no affiliation with arXiv or Cornell University. It only accesses publicly available paper metadata. Use the results in accordance with the source's terms.