CWTS Leiden Ranking Scraper avatar

CWTS Leiden Ranking Scraper

Pricing

from $5.43 / 1,000 results

Go to Apify Store
CWTS Leiden Ranking Scraper

CWTS Leiden Ranking Scraper

Scrapes CWTS Leiden Ranking university data by publication period and optional country filter. Returns each university as a flat row with impact, collaboration, and open access indicators.

Pricing

from $5.43 / 1,000 results

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

6 hours ago

Last modified

Share

ParseForge

CWTS Leiden Ranking Scraper

Scrape CWTS Leiden Ranking university data for any publication period and country, up to a million rows per run. Each university comes with its scientific impact indicators, collaboration metrics, and open access stats. No login or API key. Export to CSV, JSON, Excel, or XML.

The CWTS Leiden Ranking is a leading source for university research performance, but its official website offers no bulk export and manual copying is slow. This scraper reads the public ranking pages directly, filtered by publication period and country, and returns each university in one fixed schema. It covers scientific impact, collaboration, and open access indicators for hundreds of universities worldwide.

Who uses itWhat they scrape CWTS Leiden Ranking for
University administratorsBenchmark their institution against peers on research impact and collaboration
Higher education consultantsBuild comparative reports for clients on university performance
Academic researchersAnalyze trends in scientific output and open access across countries
Policy analystsTrack national research strengths and international collaboration patterns

What it does

This Actor collects CWTS Leiden Ranking university records by publication period and optional country filter, and returns each university as a flat row with its ranking metrics.

  • ๐Ÿ“Š Publication period filter: choose from 2019-2022, 2018-2021, or 2017-2020 to match your analysis window.
  • ๐ŸŒ Country filter: narrow results to a single country by name, such as Germany or Japan.
  • ๐Ÿ”ข Maximum universities: set a cap from 1 to 1,000,000 rows per run to control dataset size.
  • ๐Ÿ“ฅ Structured output: each university is returned as a flat row with all ranking indicators, ready for export.

Results export to CSV, JSON, Excel, or XML, or straight from the API.

What you can do with CWTS Leiden Ranking data

๐Ÿซ Benchmark university performance.

A university planning office runs the scraper for the latest period and their country to compare their institution's impact and collaboration scores against national peers.

๐Ÿ“ˆ Track research trends over time.

A higher education analyst collects data for multiple publication periods to see how open access rates and international collaboration have changed across universities.

๐ŸŒ Map global research strengths.

A policy researcher scrapes all universities without a country filter to identify which nations lead in scientific impact and which are rising.

๐Ÿ“Š Feed dashboards and reports.

A data engineer schedules the scraper to refresh a dashboard with the latest CWTS Leiden Ranking data for internal stakeholders.

Why choose this scraper

What you get
No API keyAccess the public ranking data without registration or rate limits
Bulk exportDownload hundreds or thousands of university records in one run
Fixed schemaEvery row has the same fields, so you can merge runs and compare periods
Up-to-dateScrapes the latest published ranking data for the selected period

How it compares

No other Store actor targets CWTS Leiden Ranking the same way, so the honest comparison is with the alternatives teams actually weigh.

CWTS Leiden Ranking ScraperBuild it in-houseBy hand
SetupRun it now, zero configDays of engineeringNone, but hours per pull
When CWTS Leiden Ranking changesMaintained for youYou fix itYou re-learn the page
Proxies, retries, anti-botBuilt inYour problemBrowser only
OutputFixed JSON schema, CSV/Excel exportWhatever you buildCopy-paste
CostPay per resultEngineering timeAnalyst hours

Configure the run

Drive the Actor with a publication period and optional country filter, and set a maximum number of universities to collect per run. The Input tab lists every parameter.

A first run with the defaults:

{
"maxItems": 10
}

A larger pull:

{
"maxItems": 200
}

Pricing

Pay-per-result: $0.006 per result collected. You pay only for the results written to your dataset.

Results collectedApproximate cost
100 results$0.60
1,000 results$6.00
10,000 results$60.00

New Apify accounts start with $5 in free credit.

Free users

Free-plan runs return up to 10 results as a preview. Upgrade your Apify plan to collect up to 1,000,000 results per run.

Run it

  1. Create a free Apify account with $5 in credit.
  2. Open the CWTS Leiden Ranking Scraper.
  3. Set your inputs and any filters, then click Start.
  4. Export the results as CSV, Excel, JSON, or XML from the Dataset tab.

Run it programmatically through the Apify API (run-sync-get-dataset-items) or the ApifyClient for JavaScript and Python.

Use with AI agents (MCP)

Give an AI agent live access to CWTS Leiden Ranking through the Model Context Protocol. Add the Actor to Claude, Cursor, or any MCP client:

$claude mcp add --transport http apify "https://mcp.apify.com?tools=parseforge/leiden-ranking-scraper"

Then prompt it in plain language to run the scraper and read back the results.

Troubleshooting

Why am I getting no results?

Make sure the country filter is spelled correctly and matches the name used on the CWTS Leiden Ranking site. Also verify that the selected publication period is valid.

The run stops before collecting all universities.

Increase the 'Maximum universities' input to a higher number, up to 1,000,000. If you still hit a limit, check your Apify plan's memory and timeout settings.

Some fields are empty in the output.

Not all universities have data for every indicator. Empty fields mean the ranking site did not provide a value for that metric.

The scraper fails with a timeout error.

Try reducing the maximum number of universities or run during off-peak hours. You can also increase the timeout in the actor's run settings.

FAQ

QuestionAnswer
What is the CWTS Leiden Ranking?It is an annual ranking of universities worldwide based on bibliometric indicators from Web of Science data, focusing on scientific impact, collaboration, and open access.
Do I need an API key or login?No. The scraper reads the public ranking pages directly, so no registration or authentication is required.
Can I filter by country?Yes, you can enter a country name in the optional country filter to return only universities from that country.
Which publication periods are available?The scraper supports 2019-2022, 2018-2021, and 2017-2020. You can select one per run.
How many universities can I scrape in one run?You can set a maximum from 1 to 1,000,000 universities. The default is 10, but you can increase it to collect the full ranking.
What data fields are returned for each university?Each row includes the university name, country, and all ranking indicators such as scientific impact, collaboration, and open access metrics. The exact fields are shown in the sample output.
In what formats can I export the data?You can export the dataset as CSV, JSON, Excel, or XML from the Apify platform.
Is the data updated automatically?The scraper fetches the current data from the CWTS Leiden Ranking website each time you run it, so you always get the latest published figures.
Can I schedule regular scrapes?Yes, you can set up a schedule in Apify to run the scraper daily, weekly, or monthly to keep your dataset fresh.
What if I get no results?Check that your country filter matches the exact name used on the ranking site, and ensure the selected publication period is available. If the problem persists, contact support.

Browse the full ParseForge collection for more scrapers.

๐Ÿ†˜ Need help? Email parseforge@protonmail.com with your run ID, your input, and what you expected.

โš ๏ธ Disclaimer. This Actor is unofficial and is not affiliated with, endorsed by, or sponsored by Centre for Science and Technology Studies, Leiden University. It collects only publicly available data. You are responsible for using the collected data in compliance with the source's terms of service and applicable data-protection laws, including GDPR, CCPA, and PIPL. Do not use it to collect personal data unlawfully.