PyPI Packages Scraper avatar

PyPI Packages Scraper

Pricing

from $12.00 / 1,000 result items

Go to Apify Store
PyPI Packages Scraper

PyPI Packages Scraper

Pull Python package data from PyPI. Returns name, version, summary, description, classifiers, license, author, project URLs (homepage, source, issues, docs), Python version requirement, dependencies, release history, last upload, and total release count. Direct lookup by package name.

Pricing

from $12.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

6 days ago

Last modified

Share

ParseForge Banner

๐Ÿ PyPI Python Package Scraper

๐Ÿš€ Pull PyPI packages with version, license, classifiers, dependencies, vulnerabilities, release files (wheel + sdist), funding URL, and 33 fields.

The PyPI Python Package Scraper pulls rich package metadata from the Python Package Index. Output includes name, version, summary, description (truncated), license + license expression + license files, author + email, maintainer email, homepage, repository, bug tracker, docs URL, changelog URL, funding URL, classifiers, runtime dependencies, provides_extra optionals, python version requirement, yanked flag, total releases, release files (wheel + sdist with size + SHA-256 + Python version), and security vulnerabilities published by the PyPI Security team.

Direct lookup only - feed a list of package names, get rich records back. The Actor uses the JSON detail endpoint, which is the canonical source for PyPI metadata.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
Python developers, security teams, SBOM builders, ML researchers, package-discovery tools, OSS analyticsPython supply chain analysis, vulnerability tracking, SBOM generation, dependency-graph extraction, ecosystem health monitoring

๐Ÿ“‹ What the PyPI Python Package Scraper does

Five filtering workflows in a single run:

  • ๐Ÿ†” Direct lookup. One package per line, plain names.
  • ๐Ÿšจ Vulnerabilities included. Security advisories from the PyPI Security DB.
  • ๐Ÿ“ฆ Release files. Per-version wheel + sdist with sizes and SHA-256.
  • โš–๏ธ License + license files. Standard license, license expression (PEP 639), license file paths.
  • ๐Ÿ”— Project URLs. Homepage, repo, bugs, docs, changelog, funding, all in one map.

๐Ÿ’ก Why it matters: clean, server-side filtering and fresh data on every run.

๐Ÿ“Š Data fields

Each record includes: authorName, homepage, lastUpload, license, name, pypiUrl, pythonRequires, summary, totalReleases, version. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Find the PyPI Python Package Scraper on the Apify Store.
  3. ๐ŸŽฏ Set input. Pick filters and maxItems.
  4. ๐Ÿš€ Run it. Click Start.
  5. ๐Ÿ“ฅ Download. Grab results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.

โš ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the Python Software Foundation, the PyPI maintainers, or any individual package author. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.