arXiv Papers Scraper – Search Research Papers & Abstracts
Pricing
from $2.00 / 1,000 papers
arXiv Papers Scraper – Search Research Papers & Abstracts
Search arXiv research papers by keywords, categories (cs.AI, cs.CL, stat.ML …) and submission date, or fetch papers by ID: title, abstract, authors, categories, dates, PDF link, DOI and journal reference. Official arXiv API, CC0 metadata. Pay only per paper returned.
Pricing
from $2.00 / 1,000 papers
Rating
0.0
(0)
Developer
Steven Kramp
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Categories
Share
Search arXiv research papers by keywords, categories and submission date, or fetch specific papers by ID. You get clean, structured metadata with the full abstract, authors, categories, dates, PDF link, DOI and journal reference. The data comes from arXiv's official API, and arXiv metadata is free to reuse (CC0).
$2 per 1,000 papers. Up to 10,000 papers per run. You only pay for papers returned.
What you get for each paper
| Field | Example |
|---|---|
arxivId, version, absUrl, pdfUrl | 1706.03762 · v7 |
title, abstract | Attention Is All You Need · full abstract |
authors | list of author names |
primaryCategory, categories | cs.CL · cs.CL, cs.LG |
published, updated | submission and last update date |
doi, journalRef, comment | when the authors provided them |
query, position, scrapedAt | run context |
Use cases
- Research monitoring: get every new paper in your field each day or week with a schedule ("cs.AI", "llm agents").
- Literature reviews: pull hundreds of papers with abstracts into a spreadsheet or reference manager.
- AI and RAG pipelines: feed fresh abstracts and PDF links into your own models and knowledge bases.
- Trend analysis: count papers per topic and month to see which research areas grow.
How to use
{"searchQuery": "llm agents","categories": ["cs.AI", "cs.CL"],"dateFrom": "2026-01-01","sortBy": "newest","maxResults": 500}
Advanced arXiv syntax works too, e.g. ti:transformer AND au:vaswani. To get specific papers, list IDs or URLs in arxivIds.
Pricing
Pay per event: $0.002 per paper returned ($2 per 1,000). Platform usage is included.
Use with AI agents (MCP)
This Actor works as a tool for AI assistants and agents – Claude, ChatGPT, Cursor, VS Code, n8n and other MCP clients – through Apify's hosted MCP server. Add this server URL to your client:
https://mcp.apify.com?tools=stevenkramp/arxiv-papers-scraper
Sign in with your Apify account when asked. Your agent can then call the Actor in plain language, for example: "Find this week's new arXiv papers on AI agents and summarize the five most interesting abstracts." Runs started by your agent are normal Actor runs on your Apify account at the same pay-per-event price.
More from stevenkramp
Other Actors by the same developer – same quality standards, pay only for results:
Search & trends
- Google Trends Scraper – interest over time, regions, rising queries
- Keyword Trends Finder – keyword ideas with trend direction
- Google News Scraper – news articles with real URLs
- Google Images Scraper – full-size image URLs
- Google Shopping Scraper – prices and merchants
- Google Jobs Scraper – job listings
Apps
- Google Play Store Scraper – Android app data and rankings
- Google Play Reviews Scraper – Play Store reviews and ratings
- Apple App Store Scraper – iPhone, iPad and Mac app data
- Shopify App Store Scraper – Shopify apps and pricing plans
Research & media
- Apple Podcasts Scraper – podcasts with latest episodes
Websites & places
- Website SEO Audit – broken links, titles, sitemap
- Germany Neighborhood Profile – German neighborhood rents and vacancy
FAQ
How fast is it? arXiv asks for one request every 3 seconds, and we follow this. One request returns up to 200 papers, so 1,000 papers take about 15 seconds.
Full texts? We return the abstract and the PDF link. The PDFs are hosted by arXiv. Please respect each paper's license when you reuse full texts.
Something broken? Open an issue in the Issues tab. We fix problems quickly.
Thank you to arXiv for use of its open access interoperability.