arXiv API — preprint search (AI, physics, math…) no key avatar

arXiv API — preprint search (AI, physics, math…) no key

Pricing

from $1.50 / 1,000 arxiv queries

Go to Apify Store
arXiv API — preprint search (AI, physics, math…) no key

arXiv API — preprint search (AI, physics, math…) no key

Search arXiv preprints by query, category and author (sorted by relevance or newest) and fetch a paper by id with its abstract, authors, categories, DOI and PDF link. Official arXiv API. Clean HTTP API + MCP, no key. Pay per query.

Pricing

from $1.50 / 1,000 arxiv queries

Rating

0.0

(0)

Developer

Synthetic

Synthetic

Maintained by Community

Actor stats

0

Bookmarked

1

Total users

0

Monthly active users

7 days ago

Last modified

Categories

Share

Search arXiv — the open archive of 2.4M+ preprints in AI/ML, physics, mathematics, quantitative biology, statistics and more — behind a clean HTTP API (and MCP tool). Search by keywords, category and author, sort by relevance or newest, and fetch a paper by id with its abstract, authors, categories, DOI, journal reference and PDF link. Built for research assistants, paper trackers, literature tools and AI agents. No API key. Pay per query.

  • 🔎 Search — keywords + arXiv category (e.g. cs.AI, cs.CL, stat.ML) + author; relevance or most-recent.
  • 📄 Paper — id → title, full abstract, authors, primary + all categories, DOI, journal ref, comment, abs & PDF URLs.
  • 🆕 Track the latest — sort by newest to watch a field or category.
  • 🧠 Perfect for agents — arXiv is where AI research lives; give your agent real paper search.
  • 💵 Pay per query · no key.

The use case it was built for

"What are the newest cs.CL papers on retrieval-augmented generation?" or "get me the abstract and PDF for 2005.11401" — clean JSON in one call, no scraping, no key. Ideal for RAG pipelines and research agents.

Sample output (search, query "large language models", category "cs.CL", sort recent)

{
"ok": true, "query": "large language models", "category": "cs.CL",
"totalResults": 48210, "count": 1,
"papers": [
{ "arxivId": "2609.20817v1", "title": "…", "authors": ["…"],
"primaryCategory": "cs.CL", "categories": ["cs.CL", "cs.AI"],
"published": "2026-09-17T…", "summary": "…",
"absUrl": "https://arxiv.org/abs/2609.20817", "pdfUrl": "https://arxiv.org/pdf/2609.20817" }
]
}

How to use

  • HTTP API (Standby): POST or GET /search, /paper. GET / returns help. Example: GET /search?query=diffusion&category=cs.CV&sort=recent.
  • Normal run: pass { operation, query | id, ... }; results land in the dataset + key-value store.
  • MCP tool: expose it to your agent via Apify's MCP server.

Input

  • operation — search · paper.
  • query — keywords. category — arXiv category. author — author name. sort — relevance / recent. limit — cap papers.
  • id — for paper (e.g. 2301.00001; a full arxiv.org URL works too).

Pricing

Per query from $0.004 (down to $0.001 on higher tiers). One call = one query (search returns many papers at once).

  • Crossref API — scholarly works and DOI metadata across all publishers.
  • PubMed API — biomedical literature search.

FAQ

Do I need a key? No — the arXiv API is keyless.

Which categories can I use? Any arXiv category code, e.g. cs.AI, cs.LG, cs.CL, cs.CV, stat.ML, math.OC, physics.optics, q-bio.NC.

Is the abstract included? Yes — arXiv always provides the abstract (summary).

Notes & limits

Data from arXiv (export.arxiv.org). Up to 50 papers per search call. arXiv asks clients to be gentle; this actor retries with backoff. Thank you to arXiv for open access to its e-prints.