Primary Subject
primary_subject
Optional
Primary arXiv subject/category, e.g. Computation and Language (cs.CL)
ArXiv Paper Metadata Scraper
Pricing
$1.00 / 1,000 paper scrapeds
Extracts full metadata (title, authors, abstract, subjects, dates, DOI, PDF/HTML/source links) from arXiv paper abstract pages given a list of URLs.
Url
url
Optional
The arXiv abstract page URL that was scraped
Arxiv Id
arxiv_id
Optional
The arXiv identifier, e.g. 1706.03762
Title
title
Optional
Paper title
Abstract
abstract
Optional
Full abstract text
Primary Subject
primary_subject
Optional
Primary arXiv subject/category, e.g. Computation and Language (cs.CL)
Subjects
subjects
Optional
Semicolon-separated list of all listed subjects/categories
Submitted Date
submitted_date
Optional
Original submission date
Last Revised Date
last_revised_date
Optional
Date of the most recent revision
Latest Version
latest_version
Optional
Latest version label, e.g. v7
Num Versions
num_versions
Optional
Total number of versions submitted
Pdf Url
pdf_url
Optional
Direct link to the PDF
Html Url
html_url
Optional
Link to the experimental HTML rendering, if available
Source Url
source_url
Optional
Link to the TeX/source archive
Doi
doi
Optional
arXiv-issued DOI, if present
License Url
license_url
Optional
License URL for the submission
Comments
comments
Optional
Author comments field, e.g. page/figure count