# Bluesky Posts Scraper (`blueskyscraper/bluesky-posts-scraper`) Actor

Scrape Bluesky posts by keyword, hashtag or phrase, from any account or custom feed: text, author, date, likes, reposts, replies, quotes, images, video, links and hashtags. Latest search goes back in time; language, date, domain and engagement filters. No login.

- **URL**: https://apify.com/blueskyscraper/bluesky-posts-scraper.md
- **Developed by:** [BlueskyScraper](https://apify.com/blueskyscraper) (community)
- **Categories:** Social media, Agents, MCP servers
- **Stats:** 1 total users, 0 monthly users, 0.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.70 / 1,000 posts

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

**Bluesky Posts Scraper** collects public Bluesky posts as clean rows without a login: **search by keyword, hashtag or exact phrase** that pages back in time, **all posts of any account**, the posts of any **custom feed**, and the **quote posts** of a post. Every row has text, author, date, likes, reposts, replies, quotes, images, video, link card, hashtags, mentions and the reply context. **$1 per 1,000 posts, no start fee, error rows free.**

### Tested head-to-head — 28 September 2026

![Bluesky scrapers tested head-to-head: this Actor returns search results in 2.8 s, 300 of 300 posts, 43 fields per post, follower counts included, $1 per 1,000 posts and $0.50 per 1,000 followers](https://api.apify.com/v2/key-value-stores/tZJ3TnhQJ1jqUQVXq/records/bluesky-tested-2026-09-28.png)

The same requests went to this Actor and to the six most-used Bluesky scrapers on the Apify Store: 20 latest posts for `coffee`, 20 posts of @nytimes.com, 20 followers of @nytimes.com, and a 300-post deep search.

| | This Actor | The six others |
|---|---|---|
| Keyword search, 20 posts | **2.8 s** | 3.3–28.9 s |
| 300 posts asked for one keyword | **300** | 50–300 (three tested) |
| Fields in a post row | **43** | 18–32 |
| Followers export with each person's follower count | **yes, at no extra cost** | 2 of 3 tested |
| Price per 1,000 followers | **$0.50** | $1.25–3.00 |
| Price per 1,000 posts | **$1.00** | $1.00–3.00 |
| Start fee per run | **none** | 3 of 6 charge $0.004–0.005 |

### What is Bluesky Posts Scraper?

A **Bluesky search API alternative** for social listening, brand monitoring, trend research, sentiment analysis and AI datasets. Bluesky refuses to page search results for visitors who are not logged in; this Actor pages **Latest** results back by date instead, so one keyword can return thousands of posts. It needs no account, no app password and no cookies.

| Mode | You give | You get |
|---|---|---|
| **Search** | keywords, #hashtags, "exact phrases", search operators | posts, Latest paged back in time or Top ranked (first 100) |
| **Posts** | handles or profile URLs | the account's posts, own threads, replies or media posts |
| **Custom feed** | feed URL in *Bluesky URLs* | the posts the feed shows |
| **Quotes** | post URLs | posts quoting the post |

Search filters: **language**, **date window**, **only posts by** an account, **only posts mentioning** an account, **linking to a domain** or an exact URL, **must have hashtags**, **minimum likes / reposts / replies**, and **only posts that contain the keyword**.

### What data do you get?

`url`, `text`, `createdAt`, `language`, `handle`, `displayName`, `did`, `authorVerified`, `likeCount`, `repostCount`, `replyCount`, `quoteCount`, `engagement`, `mediaType`, `images[]` (full size, alt text), `video` (playlist, thumbnail), `externalLink` (url, title, description), `hashtags`, `mentions`, `links`, `isReply`, `replyToUrl`, `rootUrl`, `isQuote`, `quotedPost`, `isRepost`, `isPinned`, `labels`, `containsKeyword`, `query`, `source`, `scrapedAt`.

### How much does it cost?

| Row | Price |
|---|---|
| **Post** | **$0.001** — $1 per 1,000 posts |

No start fee, no proxy charge, no browser — and error rows are free. **Maximum rows per run** is a hard ceiling on the bill.

### How to use it

1. Choose **Search** and type keywords, hashtags or phrases — or choose **Posts** and paste handles.
2. Set **Results per account, keyword or URL** and, for search, **Latest** or **Top**.
3. Add filters, click **Start**, download JSON, CSV or Excel, or read the rows by API.

### ⬇️ Input

```json
{
    "mode": "search",
    "keywords": ["coffee", "#specialtycoffee"],
    "sort": "latest",
    "languages": ["en"],
    "minLikes": 1,
    "maxResultsPerInput": 1000
}
```

### ⬆️ Output

```json
{
    "type": "post",
    "url": "https://bsky.app/profile/bsky.app/post/3l6oveex3ii2l",
    "text": "👋  Bluesky is an open social network that gives creators independence from platforms…",
    "createdAt": "2024-10-17T07:06:51.491Z",
    "handle": "bsky.app",
    "likeCount": 63700,
    "repostCount": 9533,
    "replyCount": 8596,
    "quoteCount": 708,
    "mediaType": "text",
    "language": "en",
    "source": "posts"
}
```

### Use cases

- **Brand and competitor monitoring** — mentions, keywords and links to your domain, scheduled hourly with **Monitoring mode** so each run returns only new posts.
- **Trend and news research** — what a community is posting about a topic right now, filtered by language and engagement.
- **AI and NLP datasets** — clean text with language tags, stable IDs and timestamps.

### ❓ FAQ

#### Is it legal to scrape Bluesky?

It reads only public posts Bluesky shows anyone without a login, and skips accounts that asked not to be shown to logged-out visitors. Posts can contain personal data protected by laws like the GDPR — have a lawful reason to process it.

#### Why does Top stop at 100 posts?

That is all Bluesky shows a visitor who is not logged in. **Latest** goes back as far as you ask.

#### Can I get only new posts on each run?

Yes — set **Monitoring mode — store name** and schedule the Actor.

### You might also like

| Actor | What it does |
|---|---|
| [Bluesky Scraper](https://apify.com/blueskyscraper/bluesky-scraper) | All twelve modes in one Actor |
| [Bluesky Profile Scraper](https://apify.com/blueskyscraper/bluesky-profile-scraper) | Profiles with follower counts, bio e-mails and links |
| [Bluesky Followers Scraper](https://apify.com/blueskyscraper/bluesky-followers-scraper) | Followers, following, likers, reposters, starter pack members |
| [Bluesky Comments Scraper](https://apify.com/blueskyscraper/bluesky-comments-scraper) | Every reply under a post, plus quote posts |

# Actor input Schema

## `mode` (type: `string`):

<b>Posts</b> — an account's own posts. <b>Search</b> — public posts for a keyword, hashtag or phrase. <b>Profiles</b> — followers count, bio, links, emails. <b>Followers / Following</b> — an account's audience. <b>Comments</b> — a post and every reply under it. <b>Quotes / Likers / Reposters</b> — who reacted to a post. <b>Account search</b> — people and brands by keyword. <b>Custom feed</b> and <b>List or starter pack</b> — paste the URL into <i>Bluesky URLs</i>.

## `handles` (type: `array`):

For Posts, Profiles, Followers and Following. Handle, @handle, profile URL or DID: <code>bsky.app</code>, <code>@nytimes.com</code>, <code>https://bsky.app/profile/jay.bsky.team</code>, <code>alice</code> (read as alice.bsky.social).

## `keywords` (type: `array`):

For Search and Account search. <code>coffee</code>, <code>#bookstodon</code>, <code>"climate policy"</code> (exact phrase). Bluesky search operators work too: <code>from:nytimes.com</code>, <code>lang:de</code>, <code>domain:github.com</code>.

## `postUrls` (type: `array`):

For Comments, Quotes, Likers and Reposters: <code>https://bsky.app/profile/bsky.app/post/3l6oveex3ii2l</code>. An <code>at://</code> URI works too.

## `startUrls` (type: `array`):

Any bsky.app link, routed automatically: profiles, posts, <b>custom feeds</b> (<code>…/profile/\<handle>/feed/\<name></code>), <b>lists</b> (<code>…/lists/\<id></code>), <b>starter packs</b> (<code>bsky.app/starter-pack/…</code>), search and hashtag pages.

## `maxResultsPerInput` (type: `integer`):

Cap for each input, not for the whole run. Search pages back in time for Latest; Top returns up to 100 posts per keyword (all Bluesky shows a visitor who is not logged in).

## `sort` (type: `string`):

Search: <b>Latest</b> goes back as far as you ask; <b>Top</b> is Bluesky's ranking, first 100 posts per keyword.

## `postedAfter` (type: `string`):

<code>YYYY-MM-DD</code>. Older posts are dropped; Posts and Search stop paging once they reach them.

## `postedBefore` (type: `string`):

<code>YYYY-MM-DD</code>. With <i>Only posts after</i> this gives a date window.

## `languages` (type: `array`):

Keep only posts written in these languages, ISO codes: <code>en</code>, <code>de</code>, <code>ja</code>, <code>pt</code>. Bluesky apps tag the language of almost every post. Empty = all.

## `hashtags` (type: `array`):

Search: only posts carrying all of these hashtags (without #).

## `fromAuthor` (type: `string`):

Search: limit results to one account's posts.

## `mentions` (type: `string`):

Search: posts that mention this account — brand monitoring in one field.

## `domain` (type: `string`):

Search: posts that link to a website, e.g. <code>nytimes.com</code> — who shares your content.

## `linkUrl` (type: `string`):

Search: posts sharing one exact page.

## `strictKeywordMatch` (type: `boolean`):

Search: Bluesky also returns posts that match a keyword by stem or by a link preview. On keeps only posts whose text, hashtags or link title contain it. Every row carries <code>containsKeyword</code> either way.

## `filterKeywords` (type: `array`):

Any mode: a post is kept when its text, hashtags or link title contain any of these words. Empty = keep everything.

## `minLikes` (type: `integer`):

Drop posts with fewer likes.

## `minReposts` (type: `integer`):

Drop posts with fewer reposts.

## `minReplies` (type: `integer`):

Drop posts with fewer replies.

## `authorFeedFilter` (type: `string`):

Posts mode: the same tabs a profile shows.

## `includeReposts` (type: `boolean`):

Posts mode: keep posts the account reposted (marked <code>isRepost</code>, with <code>repostedAt</code>).

## `includeProfile` (type: `boolean`):

Posts mode: also save one profile row per account — followers, bio, links, emails.

## `maxItems` (type: `integer`):

Stop after this many rows in the whole run — a hard ceiling on cost.

## `monitoringStoreName` (type: `string`):

Give the run a name (e.g. <code>brand-watch</code>) and every later run with the same name returns <b>only posts it has not delivered before</b> — schedule it hourly and get a clean feed of new posts, nothing billed twice. Empty = normal run.

## Actor input object example

```json
{
  "mode": "search",
  "keywords": [
    "coffee"
  ],
  "maxResultsPerInput": 50,
  "sort": "latest",
  "strictKeywordMatch": false,
  "minLikes": 0,
  "minReposts": 0,
  "minReplies": 0,
  "authorFeedFilter": "posts_and_author_threads",
  "includeReposts": true,
  "includeProfile": false,
  "maxItems": 10000
}
```

# Actor output Schema

## `rows` (type: `string`):

One row per post, profile or account; error rows explain inputs that could not be read.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "keywords": [
        "coffee"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("blueskyscraper/bluesky-posts-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "keywords": ["coffee"] }

# Run the Actor and wait for it to finish
run = client.actor("blueskyscraper/bluesky-posts-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "keywords": [
    "coffee"
  ]
}' |
apify call blueskyscraper/bluesky-posts-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,blueskyscraper/bluesky-posts-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/oTI8opgRzczfMcfWI/builds/Vg0WdgDMc3WgqbGYW/openapi.json
