Maven Central Scraper | Java Package Metadata avatar

Maven Central Scraper | Java Package Metadata

Pricing

from $19.00 / 1,000 results

Go to Apify Store
Maven Central Scraper | Java Package Metadata

Maven Central Scraper | Java Package Metadata

Extract Java and Kotlin artifacts from Maven Central including group ID, artifact ID, version history, dependencies, publisher, packaging, and license info. Audit JVM dependencies, track ecosystem trends, or feed developer security, SBOM, and intelligence tools at scale.

Pricing

from $19.00 / 1,000 results

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

2 hours ago

Last modified

Share

ParseForge Banner

โ˜• Maven Central Repository Scraper

๐Ÿš€ Export Java and JVM packages from Maven Central in seconds. Search by keyword, filter by groupId, and download artifact metadata - groupId, artifactId, version, packaging, version history, and timestamps - without touching a browser or writing a single parser.

The Maven Central Scraper queries the official Maven Central Solr search API and returns structured metadata for every matching Java and JVM package. Each record includes the full Maven coordinates (groupId, artifactId, latestVersion), packaging type, version count, repository origin, and last-updated timestamp.

Maven Central is the primary public repository for JVM ecosystem packages - the definitive source for Spring, Jackson, Hibernate, Guava, Apache Commons, and hundreds of thousands of other open-source libraries. This Actor makes the entire catalog searchable and downloadable as JSON, CSV, or Excel without any setup.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
Java developers, security teams, DevOps engineers, data analysts, OSS researchers, enterprise architectsDependency auditing, supply chain analysis, ecosystem research, package discovery, build tool integration, license compliance

๐Ÿ“‹ What the Maven Central Scraper does

Five search workflows in a single run:

  • ๐Ÿ” Keyword search. Find all packages matching a library name, technology, or concept (e.g. spring, jackson, logging, kafka).
  • ๐Ÿท GroupId filter. Restrict to a specific Maven group like org.springframework, com.fasterxml.jackson, or org.apache.commons.
  • ๐Ÿ”€ Combined search. Search by keyword within a specific groupId for targeted results.
  • ๐Ÿ“ฆ Full catalog browse. Leave both fields empty to iterate through Maven Central's entire public artifact index.
  • ๐Ÿ“Š Version history. Every record includes versionCount - the total number of published versions for that artifact.

Each record includes Maven coordinates, latest version, packaging type (jar, pom, bundle, aar), version count, repository ID, and the ISO timestamp of the last published version.

๐Ÿ“Š Data fields

Each record includes: artifactId, groupId, lastUpdated, latestVersion, packaging, repositoryId, results, scrapedAt, url, versionCount. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. Create a free account w/ $5 credit on Apify.
  2. Open the Maven Central Scraper Actor page.
  3. Enter a searchQuery (e.g. spring) or groupId (e.g. org.springframework).
  4. Set maxItems - free plan gives 10, paid plan up to 1,000,000.
  5. Click Run and wait a few seconds.
  6. Download your dataset as JSON, CSV, Excel, or XML.
ActorDescription
PyPI ScraperScrape Python packages from the PyPI registry
NPM Registry ScraperExtract JavaScript packages from the npm registry
NuGet ScraperDownload .NET package metadata from NuGet.org
Crates.io ScraperScrape Rust packages from crates.io
GitHub Trending ScraperTrack trending repositories across programming languages

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for scrapers covering package registries, developer tools, job boards, and public datasets.

This Actor queries Maven Central's public search API. All data is publicly available at search.maven.org. This tool is intended for lawful research, analysis, and development purposes only.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.