GitLab Public Projects Scraper | Stars, Forks, Topics
Pricing
from $19.00 / 1,000 results
GitLab Public Projects Scraper | Stars, Forks, Topics
Harvest records from multiple Gitlab sources in a single run and get a unified, normalized result set. Pull names, identifiers, dates, descriptions, status flags and source links per record. Perfect for research, lead generation and intelligence pipelines.
Pricing
from $19.00 / 1,000 results
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
6 days ago
Last modified
Categories
Share

๐ฆ GitLab Public Projects Scraper
๐ Pull public GitLab projects with stars, forks, topics, and owners in seconds. Built on the official GitLab REST API.
The GitLab Public Projects Scraper queries the official gitlab.com/api/v4/projects endpoint and returns one normalized record per public project. Useful for tracking open-source DevOps tooling, discovering self-hosted alternatives to GitHub repos, monitoring topic communities (kubernetes, gitops, AI), or building competitive intelligence on the open-source ecosystem.
Coverage: every public project on gitlab.com. Filters by search query, topic, and sort order (stars, last activity, created date). Up to 1,000,000 records per run on the paid plan.
| ๐ฏ Target Audience | ๐ก Primary Use Cases |
|---|---|
| Developer relations | Map ecosystem of related OSS projects |
| Security teams | Find packages by topic or maintainer |
| Recruiters | Spot top GitLab contributors |
| Researchers | Study OSS contribution patterns |
๐ What the GitLab Public Projects Scraper does
- Queries the public GitLab REST API directly
- Returns 25 normalized fields per project (name, path, stars, forks, topics, license, owner...)
- Supports search, sort, topic, and ascending/descending order
- Outputs to multiple table outputs via Apify dataset
- Auto-limits to 10 items on the free plan; up to 1,000,000 on paid
๐ก Why it matters: GitLab hosts millions of projects but no public search UI exposes them at scale. This Actor turns that catalog into a queryable dataset.
๐ Data fields
Each record includes: description, id, imageUrl, name, owner, results, scrapedAt, topics, url. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
โ ๏ธ Good to Know: Without a token, GitLab's API rate-limits to ~10 requests per second per IP. The Actor paces requests automatically.
๐ How to use
- Create a free Apify account w/ $5 credit
- Open the GitLab Public Projects Scraper actor page
- Set
search,orderBy, optionaltopic, andmaxItems - Click Start and wait for the run to finish
- Use the dataset as multiple table outputs
๐ Recommended Actors
| Actor | What it does |
|---|---|
| GitHub Trending Scraper | Daily trending repos |
| Hacker News Scraper | Top tech stories |
| npm Packages Scraper | npm metadata |
| Mastodon Trends Scraper | Fediverse trends |
๐ก Pro Tip: browse the complete ParseForge collection.
โ ๏ธ Disclaimer: independent tool, not affiliated with GitLab Inc. Only publicly available data is collected.
๐ Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.