# Beehiiv Newsletter Scraper - Low-cost 💲🔥📰📬 (`delectable_incubator/beehiiv-newsletter-scraper-low-cost`) Actor

📰 Scrape Beehiiv newsletter articles from one or multiple Beehiiv publications. Extract article titles, URLs, publication dates, authors, descriptions, featured images, and other metadata. Perfect for content monitoring, newsletter analytics, market research, AI datasets and content automation 🚀📊

- **URL**: https://apify.com/delectable\_incubator/beehiiv-newsletter-scraper-low-cost.md
- **Developed by:** [Prime Scrape](https://apify.com/delectable_incubator) (community)
- **Categories:** News, SEO tools, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.00005 / actor start

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are a software tools running on the Apify platform, for all kinds of web data extraction and automation use cases.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

In JavaScript/TypeScript projects, use official [JavaScript/TypeScript client](https://docs.apify.com/api/client/js/docs.md):

```bash
npm install apify-client
```

In Python projects, use official [Python client library](https://docs.apify.com/api/client/python/docs.md):

```bash
pip install apify-client
```

In shell scripts, use [Apify CLI](https://docs.apify.com/cli/docs.md):

````bash
# MacOS / Linux
curl -fsSL https://apify.com/install-cli.sh | bash
# Windows
irm https://apify.com/install-cli.ps1 | iex
```bash

In AI frameworks, you might use the [Apify MCP server](https://docs.apify.com/integrations/mcp.md).

If your project is in a different language, use the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).


# README

## 📰 Beehiiv Newsletter Scraper 🚀📬

The Beehiiv Newsletter Scraper is a fast, reliable, and scalable tool that extracts newsletter articles from Beehiiv-powered publications.

Simply provide one or multiple Beehiiv newsletter URLs, and the scraper automatically discovers all available newsletter posts through the public sitemap before extracting detailed article data.

Perfect for AI datasets, content monitoring, market research, competitor analysis, trend tracking, newsletter archiving, and SEO research.

---

## 🔍 What This Scraper Does

Simply provide one or multiple Beehiiv newsletter URLs, and the scraper will:

✅ Automatically discover newsletter articles using the public sitemap

✅ Support bulk newsletter URLs

✅ Optionally filter articles by keywords

✅ Automatically handle pagination and discovery

✅ Extract clean, structured article data

✅ Export analysis-ready datasets

---

## 📊 Data Extracted

| Field | Description |
|------|-------------|
| `title` | Article title |
| `url` | Article URL |
| `slug` | Article slug |
| `publishedAt` | Publication date |
| `author` | Author name |
| `description` | Article description |
| `content` | Full article content |
| `readingTime` | Estimated reading time |
| `coverImage` | Featured image |
| `tags` | Article tags (if available) |
| `newsletter` | Newsletter name |
| `sourceUrl` | Source Beehiiv website |
| `scrapedAt` | Extraction timestamp |

---

## 🛠 How to Use

#### 1️⃣ Enter one or multiple Beehiiv newsletter URLs

Example:

````

{
"urls": \[
"https://www.therundown.ai",
"https://www.superhuman.ai",
"https://milkroad.com"
],
"keywords": \[
"ai",
"agents"
],
"max\_items": 100
}

```

---

#### 2️⃣ Run the Actor

The scraper will automatically:

• Read each Beehiiv sitemap

• Discover newsletter articles

• Apply keyword filtering (optional)

• Extract structured article data

• Stop once the maximum number of articles has been reached

---

#### 3️⃣ Export your dataset

Supported formats:

✅ JSON

✅ CSV

✅ Excel

✅ XML

✅ HTML

---

## 💰 Pricing

This scraper runs on a pay-per-result pricing model.

You only pay for successfully extracted articles.

💳 **Price:** $0.99 / 1,000 results

---

## 📥 Input Configuration

### Input Example

```

{
"urls": \[
"https://www.therundown.ai",
"https://www.superhuman.ai"
],
"keywords": \[
"ai",
"crypto"
],
"max\_items": 100
}

```

### Input Fields

| Field | Type | Description |
|------|------|-------------|
| `urls` | array | One or more Beehiiv newsletter URLs |
| `keywords` | array | Optional keyword filters |
| `max_items` | number | Maximum articles to extract per newsletter |

---

## 📤 Output Example

```

{
"title": "OpenAI launches new AI agents",
"url": "https://www.therundown.ai/p/openai-launches-new-ai-agents",
"slug": "openai-launches-new-ai-agents",
"publishedAt": "2026-02-14",
"author": "The Rundown AI",
"description": "Everything you need to know about the latest OpenAI release.",
"content": "Today OpenAI announced...",
"readingTime": "6 min",
"coverImage": "https://cdn.beehiiv.com/images/article.jpg",
"tags": \[
"AI",
"OpenAI",
"Agents"
],
"newsletter": "The Rundown AI",
"sourceUrl": "https://www.therundown.ai",
"scrapedAt": "2026-03-27T18:35:12.554Z"
}

````

---

## 📊 Preconfigured Dataset Views

#### Overview

• Article titles

• Publication dates

• Authors

• Reading time

• Source newsletter

• Direct article URLs

---

## 💡 Use Cases

📰 Newsletter Monitoring

Track publications from your favorite Beehiiv newsletters.

🤖 AI Training Datasets

Build large datasets of high-quality newsletter content.

📈 Market Research

Monitor trends across industries and competitors.

🔍 SEO & Content Research

Analyze topics, publishing frequency, and keyword usage.

📊 Competitive Intelligence

Track what leading newsletters publish over time.

---

## ⚡ Why Use This Scraper?

🚀 Bulk Beehiiv newsletter scraping

📚 Automatic sitemap discovery

🔎 Optional keyword filtering

⚡ Fast extraction

📄 Clean structured datasets

🌍 Ideal for analysts, researchers, marketers, journalists, and AI developers

---

## Disclaimer

This tool is independent and is not affiliated with or endorsed by Beehiiv.

---

## 📫 Support

😊 Leave a **5-star rating ⭐⭐⭐⭐⭐** if you're satisfied!

🌪️ **Storm Scraper**

https://apify.com/scrapestorm

For feature requests, custom scraping solutions, or support, contact us via Apify.

# Actor input Schema

## `urls` (type: `array`):

One or multiple Beehiiv newsletter base URLs (one per line). The actor will automatically read the public sitemap available at {url}/sitemap.xml.

Examples:
• https://www.therundown.ai
• https://www.superhuman.ai
• https://milkroad.com
## `keywords` (type: `array`):

If provided, only articles whose URL contains at least one of these keywords will be kept (e.g., 'ai', 'crypto', 'marketing'). The search is case-insensitive. Leave empty to extract all articles.
## `max_items` (type: `integer`):

Maximum number of articles to extract for each provided newsletter URL.

## Actor input object example

```json
{
  "urls": [
    "https://www.therundown.ai",
    "https://www.superhuman.ai",
    "https://milkroad.com"
  ],
  "keywords": [],
  "max_items": 60
}
````

# Actor output Schema

## `overview` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "urls": [
        "https://www.therundown.ai",
        "https://www.superhuman.ai",
        "https://milkroad.com"
    ],
    "max_items": 60
};

// Run the Actor and wait for it to finish
const run = await client.actor("delectable_incubator/beehiiv-newsletter-scraper-low-cost").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "urls": [
        "https://www.therundown.ai",
        "https://www.superhuman.ai",
        "https://milkroad.com",
    ],
    "max_items": 60,
}

# Run the Actor and wait for it to finish
run = client.actor("delectable_incubator/beehiiv-newsletter-scraper-low-cost").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "urls": [
    "https://www.therundown.ai",
    "https://www.superhuman.ai",
    "https://milkroad.com"
  ],
  "max_items": 60
}' |
apify call delectable_incubator/beehiiv-newsletter-scraper-low-cost --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=delectable_incubator/beehiiv-newsletter-scraper-low-cost",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

```json
{
    "openapi": "3.0.1",
    "info": {
        "title": "Beehiiv Newsletter Scraper - Low-cost 💲🔥📰📬",
        "description": "📰 Scrape Beehiiv newsletter articles from one or multiple Beehiiv publications. Extract article titles, URLs, publication dates, authors, descriptions, featured images, and other metadata. Perfect for content monitoring, newsletter analytics, market research, AI datasets and content automation 🚀📊",
        "version": "0.0",
        "x-build-id": "zk2k6g9GJ3mhvnhqq"
    },
    "servers": [
        {
            "url": "https://api.apify.com/v2"
        }
    ],
    "paths": {
        "/acts/delectable_incubator~beehiiv-newsletter-scraper-low-cost/run-sync-get-dataset-items": {
            "post": {
                "operationId": "run-sync-get-dataset-items-delectable_incubator-beehiiv-newsletter-scraper-low-cost",
                "x-openai-isConsequential": false,
                "summary": "Executes an Actor, waits for its completion, and returns Actor's dataset items in response.",
                "tags": [
                    "Run Actor"
                ],
                "requestBody": {
                    "required": true,
                    "content": {
                        "application/json": {
                            "schema": {
                                "$ref": "#/components/schemas/inputSchema"
                            }
                        }
                    }
                },
                "parameters": [
                    {
                        "name": "token",
                        "in": "query",
                        "required": true,
                        "schema": {
                            "type": "string"
                        },
                        "description": "Enter your Apify token here"
                    }
                ],
                "responses": {
                    "200": {
                        "description": "OK"
                    }
                }
            }
        },
        "/acts/delectable_incubator~beehiiv-newsletter-scraper-low-cost/runs": {
            "post": {
                "operationId": "runs-sync-delectable_incubator-beehiiv-newsletter-scraper-low-cost",
                "x-openai-isConsequential": false,
                "summary": "Executes an Actor and returns information about the initiated run in response.",
                "tags": [
                    "Run Actor"
                ],
                "requestBody": {
                    "required": true,
                    "content": {
                        "application/json": {
                            "schema": {
                                "$ref": "#/components/schemas/inputSchema"
                            }
                        }
                    }
                },
                "parameters": [
                    {
                        "name": "token",
                        "in": "query",
                        "required": true,
                        "schema": {
                            "type": "string"
                        },
                        "description": "Enter your Apify token here"
                    }
                ],
                "responses": {
                    "200": {
                        "description": "OK",
                        "content": {
                            "application/json": {
                                "schema": {
                                    "$ref": "#/components/schemas/runsResponseSchema"
                                }
                            }
                        }
                    }
                }
            }
        },
        "/acts/delectable_incubator~beehiiv-newsletter-scraper-low-cost/run-sync": {
            "post": {
                "operationId": "run-sync-delectable_incubator-beehiiv-newsletter-scraper-low-cost",
                "x-openai-isConsequential": false,
                "summary": "Executes an Actor, waits for completion, and returns the OUTPUT from Key-value store in response.",
                "tags": [
                    "Run Actor"
                ],
                "requestBody": {
                    "required": true,
                    "content": {
                        "application/json": {
                            "schema": {
                                "$ref": "#/components/schemas/inputSchema"
                            }
                        }
                    }
                },
                "parameters": [
                    {
                        "name": "token",
                        "in": "query",
                        "required": true,
                        "schema": {
                            "type": "string"
                        },
                        "description": "Enter your Apify token here"
                    }
                ],
                "responses": {
                    "200": {
                        "description": "OK"
                    }
                }
            }
        }
    },
    "components": {
        "schemas": {
            "inputSchema": {
                "type": "object",
                "required": [
                    "urls"
                ],
                "properties": {
                    "urls": {
                        "title": "Newsletter URLs (Bulk) 🌐",
                        "type": "array",
                        "description": "One or multiple Beehiiv newsletter base URLs (one per line). The actor will automatically read the public sitemap available at {url}/sitemap.xml.\n\nExamples:\n• https://www.therundown.ai\n• https://www.superhuman.ai\n• https://milkroad.com",
                        "default": [
                            "https://www.therundown.ai",
                            "https://www.superhuman.ai",
                            "https://milkroad.com"
                        ],
                        "items": {
                            "type": "string"
                        }
                    },
                    "keywords": {
                        "title": "Filter by Keywords (Optional) 🔎",
                        "type": "array",
                        "description": "If provided, only articles whose URL contains at least one of these keywords will be kept (e.g., 'ai', 'crypto', 'marketing'). The search is case-insensitive. Leave empty to extract all articles.",
                        "default": [],
                        "items": {
                            "type": "string"
                        }
                    },
                    "max_items": {
                        "title": "Maximum Articles per Newsletter 🎯",
                        "type": "integer",
                        "description": "Maximum number of articles to extract for each provided newsletter URL.",
                        "default": 60
                    }
                }
            },
            "runsResponseSchema": {
                "type": "object",
                "properties": {
                    "data": {
                        "type": "object",
                        "properties": {
                            "id": {
                                "type": "string"
                            },
                            "actId": {
                                "type": "string"
                            },
                            "userId": {
                                "type": "string"
                            },
                            "startedAt": {
                                "type": "string",
                                "format": "date-time",
                                "example": "2025-01-08T00:00:00.000Z"
                            },
                            "finishedAt": {
                                "type": "string",
                                "format": "date-time",
                                "example": "2025-01-08T00:00:00.000Z"
                            },
                            "status": {
                                "type": "string",
                                "example": "READY"
                            },
                            "meta": {
                                "type": "object",
                                "properties": {
                                    "origin": {
                                        "type": "string",
                                        "example": "API"
                                    },
                                    "userAgent": {
                                        "type": "string"
                                    }
                                }
                            },
                            "stats": {
                                "type": "object",
                                "properties": {
                                    "inputBodyLen": {
                                        "type": "integer",
                                        "example": 2000
                                    },
                                    "rebootCount": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "restartCount": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "resurrectCount": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "computeUnits": {
                                        "type": "integer",
                                        "example": 0
                                    }
                                }
                            },
                            "options": {
                                "type": "object",
                                "properties": {
                                    "build": {
                                        "type": "string",
                                        "example": "latest"
                                    },
                                    "timeoutSecs": {
                                        "type": "integer",
                                        "example": 300
                                    },
                                    "memoryMbytes": {
                                        "type": "integer",
                                        "example": 1024
                                    },
                                    "diskMbytes": {
                                        "type": "integer",
                                        "example": 2048
                                    }
                                }
                            },
                            "buildId": {
                                "type": "string"
                            },
                            "defaultKeyValueStoreId": {
                                "type": "string"
                            },
                            "defaultDatasetId": {
                                "type": "string"
                            },
                            "defaultRequestQueueId": {
                                "type": "string"
                            },
                            "buildNumber": {
                                "type": "string",
                                "example": "1.0.0"
                            },
                            "containerUrl": {
                                "type": "string"
                            },
                            "usage": {
                                "type": "object",
                                "properties": {
                                    "ACTOR_COMPUTE_UNITS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATASET_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATASET_WRITES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "KEY_VALUE_STORE_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "KEY_VALUE_STORE_WRITES": {
                                        "type": "integer",
                                        "example": 1
                                    },
                                    "KEY_VALUE_STORE_LISTS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "REQUEST_QUEUE_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "REQUEST_QUEUE_WRITES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATA_TRANSFER_INTERNAL_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATA_TRANSFER_EXTERNAL_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "PROXY_RESIDENTIAL_TRANSFER_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "PROXY_SERPS": {
                                        "type": "integer",
                                        "example": 0
                                    }
                                }
                            },
                            "usageTotalUsd": {
                                "type": "number",
                                "example": 0.00005
                            },
                            "usageUsd": {
                                "type": "object",
                                "properties": {
                                    "ACTOR_COMPUTE_UNITS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATASET_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATASET_WRITES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "KEY_VALUE_STORE_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "KEY_VALUE_STORE_WRITES": {
                                        "type": "number",
                                        "example": 0.00005
                                    },
                                    "KEY_VALUE_STORE_LISTS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "REQUEST_QUEUE_READS": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "REQUEST_QUEUE_WRITES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATA_TRANSFER_INTERNAL_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "DATA_TRANSFER_EXTERNAL_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "PROXY_RESIDENTIAL_TRANSFER_GBYTES": {
                                        "type": "integer",
                                        "example": 0
                                    },
                                    "PROXY_SERPS": {
                                        "type": "integer",
                                        "example": 0
                                    }
                                }
                            }
                        }
                    }
                }
            }
        }
    }
}
```
