# GitHub Scraper — repos, trending, users (no token) (`vincentkirui/github-scraper`) Actor

Structured GitHub data with no token: search repositories, find trending repos, fetch full repo details, or pull user/org profiles and their repos. Uses the public GitHub REST API.

- **URL**: https://apify.com/vincentkirui/github-scraper.md
- **Developed by:** [Vincent Kirui](https://apify.com/vincentkirui) (community)
- **Categories:**
- **Stats:** 2 total users, 1 monthly users, 50.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.50 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## GitHub Scraper — repos, trending, users (no token)

Structured **GitHub** data with no token and no login, via the public REST API. Four modes:

| Mode | What you get | Use it for |
|---|---|---|
| **Search repos** | Repositories matching a query, sorted by stars / forks / updated | Market & competitor research, tech landscape mapping |
| **Trending** | Most-starred repos pushed in the last N days (optional language) | Spotting rising projects, daily "what's hot" digests |
| **Repo details** | Full details for any owner/repo | Enrichment, dependency/tech due-diligence |
| **Users** | A user's or org's profile + their repositories | Developer sourcing, recruiter research, OSS analytics |

### Built for AI agents

As an MCP tool, an agent can answer "what are the trending Rust repos this week?" or "get
this repo's stats" in one structured call.

### Output

Flat JSON rows. Repos: `full_name`, `description`, `stars`, `forks`, `language`, `topics`,
`license`, `created_at`, `updated_at`, `url`. Users: `login`, `name`, `bio`, `company`,
`followers`, `public_repos`, `url`.

### Input examples

```json
{ "mode": "search", "query": "llm agent", "language": "python", "minStars": 500, "maxItems": 200 }
```

```json
{ "mode": "trending", "language": "rust", "sinceDays": 7, "maxItems": 50 }
```

```json
{ "mode": "user", "users": ["torvalds"], "includeRepos": true }
```

### Notes

Uses the unauthenticated GitHub REST API (rate-limited per IP); requests route through
rotating proxies with backoff so larger pulls complete. Public data only, no login.

### Pricing

Pay-per-event: a small actor-start fee plus a per-item charge — you pay for exactly the data
you pull.

# Actor input Schema

## `mode` (type: `string`):

search = repos by query; trending = most-starred recently; repo = full repo details; user = profiles + repos

## `query` (type: `string`):

Repository search query, e.g. 'llm agent'.

## `language` (type: `string`):

Restrict to a language, e.g. python, rust, typescript.

## `minStars` (type: `integer`):

Only repos with at least this many stars.

## `sort` (type: `string`):

Sort repos by stars, forks, or last updated.

## `sinceDays` (type: `integer`):

Repos pushed in the last N days.

## `repos` (type: `array`):

owner/repo entries, e.g. openai/openai-python.

## `users` (type: `array`):

Usernames or org names.

## `includeRepos` (type: `boolean`):

Also return each user's repositories.

## `maxItems` (type: `integer`):

Stop after this many results.

## Actor input object example

```json
{
  "mode": "search",
  "query": "llm agent",
  "sort": "stars",
  "sinceDays": 7,
  "repos": [
    "torvalds/linux"
  ],
  "users": [
    "torvalds"
  ],
  "includeRepos": false,
  "maxItems": 100
}
```

# Actor output Schema

## `results` (type: `string`):

Scraped items in the default dataset.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "mode": "search",
    "query": "llm agent"
};

// Run the Actor and wait for it to finish
const run = await client.actor("vincentkirui/github-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "mode": "search",
    "query": "llm agent",
}

# Run the Actor and wait for it to finish
run = client.actor("vincentkirui/github-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "mode": "search",
  "query": "llm agent"
}' |
apify call vincentkirui/github-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,vincentkirui/github-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/FLYFHcyKLtkzQxwHu/builds/g0qI0sf2e4GZTzxcu/openapi.json
