CORE Open Research Scraper
Pricing
from $1.99 / 1,000 search results
CORE Open Research Scraper
Search CORE's public open-access research index with bounded pagination, stable normalization, deduplication, safe retries, and optional proxies.
Pricing
from $1.99 / 1,000 search results
Rating
0.0
(0)
Developer
Search API
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
8 days ago
Last modified
Categories
Share
Search CORE's public open-access works index using a topic, title, author, DOI, institution, or CORE query expression. The Actor follows the public API's offset pagination and returns a compact, stable dataset of real works.
Input example
{"query": "machine learning","maxItems": 50,"pageSize": 25,"maxPages": 2,"requestDelayMs": 300,"useApifyProxy": false}
maxResults remains a backward-compatible alias for maxItems. The Actor supports direct access, authorized Apify Proxy groups/country selection, or custom HTTP(S) proxy URLs.
Output
Each dataset row contains a stable CORE identity, title, deduplicated author names, public CORE URL, query/page/rank context, and optional abstract, DOI, type, publisher, field, dates, citations, and public download URL. Missing optional values are omitted. Failures and no-results never become placeholder work rows; the OUTPUT key contains status, page counts, total hits, bounded failure categories, connection mode, and completion time.
The Actor validates HTTP status and content type before parsing JSON, enforces a response-size bound, rejects malformed or changed payloads, retries only temporary failures, deduplicates by CORE ID, and obeys hard page/result limits. It never stores raw API payloads or logs proxy credentials, authorization headers, cookies, tokens, or sensitive response bodies. It does not bypass authentication, CAPTCHA, paywalls, regional restrictions, or other access controls.