PyPI Scraper - Python Package Search & Stats
Pricing
from $19.00 / 1,000 results
PyPI Scraper - Python Package Search & Stats
Search and scrape Python package data from PyPI including versions, authors, licenses, keywords, download stats, and classifiers. Export to CSV, Excel, JSON, XML.
Pricing
from $19.00 / 1,000 results
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
8 days ago
Last modified
Categories
Share

π PyPI Python Package Scraper
π Export Python package data from PyPI in seconds. Search 812,000+ packages by keyword and get version, author, license, keywords, download stats, classifiers, and more. No API key, no registration required.
The PyPI Scraper searches the Python Package Index and returns 15 fields per record including name, version, summary, author, license, homepage, repository URL, keywords, weekly and monthly download counts, Python version requirements, classifiers, and publish date. The underlying data comes directly from PyPI's public JSON API and is the same catalog used by pip install.
The index covers every publicly released Python package - from the most downloaded frameworks like Django, Flask, and NumPy down to niche utilities and personal projects. This Actor searches by keyword, resolves full metadata for each match, and enriches with real download statistics. Your dataset is ready to download as CSV, Excel, JSON, or XML in under a minute.
| π― Target Audience | π‘ Primary Use Cases |
|---|---|
| Data engineers, Python developers, package analysts, security researchers, DevOps teams, BI analysts | Dependency auditing, license compliance, tech-stack research, competitive analysis, package discovery |
π What the PyPI Scraper does
Five research workflows in a single run:
- π Keyword search. Find packages matching any search term across package names - "machine learning", "web scraping", "django", "data pipeline", "cli".
- π¦ Full metadata extraction. Name, version, summary, author, license, homepage, and repository URL from the PyPI JSON API.
- π Download statistics. Weekly and monthly download counts from the PyPI Stats API - instantly spot popular vs. niche packages.
- π·οΈ Classifier taxonomy. Full PyPI classifier list per package - development status, programming language versions, OS compatibility, topics.
- π Freshness signals. Last publish date and Python version requirement for every record.
π‘ Why it matters: The Python ecosystem has over 812,000 packages on PyPI. Manually auditing which libraries match a topic, what their license is, and how actively they are maintained is hours of work. This Actor delivers a structured dataset in under a minute - ready to import into Notion, Airtable, BigQuery, or your CI pipeline.
π Data fields
Each record includes: author, lastPublished, license, name, requiresPython, results, summary, totalDownloads, url, version, weeklyDownloads. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
π How to use
- Create a free account - includes $5 free credit.
- Open PyPI Python Package Scraper in the Apify Store.
- Enter your search query (e.g.
"machine learning","django","data pipeline"). - Set Max Items (free plan preview: 10, paid: up to 1,000,000).
- Click Start and wait a few seconds.
- Download your dataset as CSV, Excel, JSON, or XML.
π Recommended Actors
| Actor | What it does |
|---|---|
| npm Registry Scraper | Scrape JavaScript package metadata and download stats from the npm registry |
| Product Hunt Scraper | Extract product launches, makers, topics, and upvote counts from Product Hunt |
| GitHub Scraper | Collect repository metadata, stars, forks, and contributor data from GitHub |
| Upwork Scraper | Search Upwork job postings and freelancer profiles by keyword |
| Remotive Scraper | Collect remote tech job listings across categories and companies |
π‘ Pro Tip: browse the complete ParseForge collection for 50+ specialized data scrapers covering jobs, finance, aviation, government data, and developer tools.
Disclaimer: This Actor accesses publicly available data from PyPI's official JSON API (pypi.org/pypi/{name}/json) and the pypistats.org recent downloads API. Both APIs are open and documented by the Python Software Foundation. No login, scraping of private pages, or circumvention of access controls is involved. Use responsibly and in accordance with PyPI's terms of service.
π Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.