# Product Hunt Scraper — Daily Launches, Votes & Dataset (`kaankaan2635/producthunt-scraper`) Actor

Product Hunt dataset of daily leaderboards: launch rank, votes, topics and taglines, with product ratings and first launch dates.

- **URL**: https://apify.com/kaankaan2635/producthunt-scraper.md
- **Developed by:** [Kaan Salgır](https://apify.com/kaankaan2635) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 launch scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Product Hunt Scraper — Daily Launches, Votes & Dataset

This **Product Hunt scraper** builds a **Product Hunt dataset** from the daily leaderboards: every launch with its **rank, vote count, topics and tagline**, plus each product's rating, review count and first launch date.

No account, no API key, no rate-limited token. If you came looking for the Product Hunt API and found the quota too small, this is the alternative.

### Product Hunt API alternative: pull a date range, not a feed

Most Product Hunt scrapers hand you the current homepage. This one is built around the **daily leaderboard**, so you can pull a date range and see how launches actually performed:

```json
{
  "dateFrom": "2026-09-01",
  "dateTo": "2026-09-22",
  "topics": ["artificial-intelligence"],
  "minVotes": 100,
  "includeProductDetails": true
}
```

That gives you every AI launch that cleared 100 votes across three weeks, each with the rank it finished at. Rank and vote count together tell you whether a category is getting more competitive or less.

The same product launching twice appears once per day, so you can follow a relaunch.

**Rank is not vote order.** Product Hunt ranks by its own score, so a #1 finish regularly has fewer raw votes than the #2 below it. Both numbers are returned separately, and the gap between them is worth watching: it is where the ranking algorithm shows itself.

### Product Hunt data: what you get

**Per launch:** `leaderboardDate` · `rank` · `name` · `tagline` · `votes` · `topics` (slug and name) · `topicSlugs` · `url` · `slug`

**Per product**, with **Include product details** on: `rating` · `ratingCount` · `description` · `datePublished` (first ever launch) · `dateModified` · `operatingSystem` · `imageUrl` · `screenshotUrl` · `websiteUrl`

`datePublished` against `leaderboardDate` is how you tell a first launch from a relaunch.

Export as JSON, CSV, Excel or XML.

### How to use this Product Hunt scraper

Three ways to choose what to scrape, and they combine:

- **Date range** — `dateFrom` and `dateTo`, up to 60 days per run.
- **Specific dates** — a list of `YYYY-MM-DD` days.
- **Product slugs** — scrape products directly, skipping the leaderboards.

With nothing set, it returns yesterday's leaderboard. Today's is deliberately not the default: it is still moving, and a rank read at noon is not the rank the day ends on.

Filter by `topics` and `minVotes`.

### FAQ

**Do I need a Product Hunt API key?**
No. This reads public leaderboard and product pages, so there is no token to obtain and no API quota to run out of.

**How far back can I go?**
Leaderboards exist for every day Product Hunt has run. The scraper caps a single run at 60 days; chain runs for longer histories.

**Can it scrape makers, hunters or their emails?**
No. Product Hunt's robots.txt disallows `/@*` — the member profile space — and this scraper requests none of it. Products, votes, topics and ratings only.

**Why do some products have no rating?**
A product only carries an `aggregateRating` once it has reviews. New launches usually have none on day one, so `rating` and `ratingCount` come back null rather than zero.

**Can I track one product over time?**
Yes. Pass its slug in `productSlugs` and schedule the run; `ratingCount` moving is the signal.

**Does it scrape topic pages?**
No. The `topics` input filters leaderboard results by topic instead, which is more reliable and costs no extra requests.

**Is scraping Product Hunt legal?**
It reads publicly published pages, honours robots.txt, and collects no personal data. How you use the output is your responsibility.

### Pricing

Pay per event: one charge per launch row and one per product row. Nothing is charged for compute time or failed requests.

### Development

```bash
npm install
npm test          # offline regression tests against saved pages
npm start         # local run; put an INPUT.json in storage/key_value_stores/default/
```

# Actor input Schema

## `dateFrom` (type: `string`):

Start of the leaderboard range, <code>YYYY-MM-DD</code>. Use with <b>Date to</b>. At most 60 days per run.

## `dateTo` (type: `string`):

End of the leaderboard range, <code>YYYY-MM-DD</code>.

## `dates` (type: `array`):

Individual days as <code>YYYY-MM-DD</code>, in addition to any range above.

## `productSlugs` (type: `array`):

Scrape these products directly, for example <code>notion</code> from <code>producthunt.com/products/notion</code>.

## `topics` (type: `array`):

Keep only launches carrying one of these topic slugs, for example <code>artificial-intelligence</code>, <code>developer-tools</code>, <code>design-tools</code>.

## `minVotes` (type: `integer`):

Drop launches below this vote count.

## `includeProductDetails` (type: `boolean`):

For every launch, also fetch its product page for the rating, review count, description and first launch date. One extra request per product.

## `maxLaunches` (type: `integer`):

Hard cap across the whole run. The same product on two different days counts twice.

## `maxConcurrency` (type: `integer`):

Parallel requests. Lower this if pages come back empty.

## `maxRequestsPerMinute` (type: `integer`):

Caps the request rate against Product Hunt.

## `proxyConfiguration` (type: `object`):

Apify Proxy is recommended for multi-day runs.

## Actor input object example

```json
{
  "dateFrom": "",
  "dateTo": "",
  "dates": [],
  "productSlugs": [],
  "topics": [],
  "minVotes": 0,
  "includeProductDetails": false,
  "maxLaunches": 500,
  "maxConcurrency": 4,
  "maxRequestsPerMinute": 60,
  "proxyConfiguration": {
    "useApifyProxy": true
  }
}
```

# Actor output Schema

## `launches` (type: `string`):

One row per launch per day, plus one row per product when product details are requested.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {};

// Run the Actor and wait for it to finish
const run = await client.actor("kaankaan2635/producthunt-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {}

# Run the Actor and wait for it to finish
run = client.actor("kaankaan2635/producthunt-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{}' |
apify call kaankaan2635/producthunt-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,kaankaan2635/producthunt-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/pM4UUpNTahaPxIhEb/builds/Zp90NO4TQey7DVHxF/openapi.json
