# Ameba Blog Posts Scraper (`scrapingmonkey/ameba-blog-posts-scraper`) Actor

Export public posts from Ameba blogs with full available text, HTML, publication dates, themes, tags, and images.

- **URL**: https://apify.com/scrapingmonkey/ameba-blog-posts-scraper.md
- **Developed by:** [ScrapingMonkey](https://apify.com/scrapingmonkey) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

**Ameba Blog Posts Scraper** — Export public posts from Ameba blogs with full available text, HTML, publication dates, themes, tags, and images. Add supported inputs and start a run to get structured public data without providing a Ameba account.

Collect public posts from an Ameba lifestyle creator. Export article text, publication dates, themes and media links to build a source-linked content catalog.

| At a glance | Details |
|---|---|
| 📥 Input | Enter an Ameba blog ID or a public ameblo.jp blog URL. Each page contains up to 20 post records. |
| 📤 Output | One post per success row, with source input retained |
| 📄 Pagination | `pagesPerBlog` limits result pages per input, with a default of 1. The source may end sooner. |
| 🔐 Login required | No |
| ⚡ Processing | Up to 5 HTTP requests concurrently with automatic retries |
| 💾 Delivery | One Apify dataset view, useful flat columns and complete nested JSON |

### What the ameba blog posts scraper collects 📊

Export public posts from Ameba blogs with full available text, HTML, publication dates, themes, tags, and images.

Data can include:

- Source identity and public links
- Available id, url, ameba id, blog id, title, text, html, published at
- Original input retained with every result
- One consistent success or failed record format

### How to collect posts from Ameba 🚀

1. Enter one or more supported inputs in `inputList`.
2. Set the result-page budget for each input.
3. Start the Actor.
4. Open the **Posts** dataset view and review the individual records.
5. Export the dataset or retrieve it from your application.

```json
{
  "inputList": [
    "tsuji-nozomi"
  ],
  "pagesPerBlog": 1
}
```

`pagesPerBlog` limits result pages per input, with a default of 1. The source may end sooner.

### Ameba posts data and complete output 📦

| Field | Type | Meaning |
|---|---|---|
| `input` | string | Original submitted input. |
| `status` | string | Result status: success or failed. |
| `id` | string or null | Stable identifier of this result. |
| `url` | string or null | Public result URL. |
| `ameba_id` | string or null | Ameba blog ID / username. |
| `blog_id` | string or null | Numeric Ameba blog ID. |
| `title` | string or null | Public blog or post title. |
| `text` | string or null | Available article or comment text with markup removed. |
| `html` | string or null | Available article or comment HTML. |
| `published_at` | string or null | Publication timestamp. |
| `updated_at` | string or null | Last update timestamp when provided. |
| `theme_id` | string or null | Ameba theme/category ID. |
| `theme_name` | string or null | Public theme/category name. |
| `thumbnail_url` | string or null | Article preview image URL. |
| `tags` | array or null | Public tags. |
| `images` | array or null | Article image URLs with available alt text and dimensions. |
| `images[].url` | string or null | Public source URL. |
| `images[].alt` | string or null | Alt when exposed by the public response. |
| `images[].width` | integer or null | Image width in pixels when supplied. |
| `images[].height` | integer or null | Image height in pixels when supplied. |
| `embed_urls` | array or null | Video, audio, or iframe URLs found in the public article HTML. |
| `is_reblog` | boolean or null | Whether the source labels the post as a reblog. |
| `source_blog_id` | string or null | Numeric source blog ID. |
| `source_post_id` | string or null | ID of the source article. |
| `is_member_only` | boolean or null | Whether this public list record refers to an approved-members-only article. |

Nested arrays and media references stay in their parent record. Downloading source media files is not part of this Actor.

Every top-level output field appears in these complete examples. The success example is normalized from a public response. Long strings and arrays are shortened here; the Actor keeps the complete available values. Values are a snapshot and can change. Nested arrays remain inside their parent result.

Representative success result with every output field (long text and arrays shortened for readability):

```json
{
  "input": "tsuji-nozomi",
  "status": "success",
  "id": "12975182730",
  "url": "https://ameblo.jp/tsuji-nozomi/entry-12975182730.html",
  "ameba_id": "tsuji-nozomi",
  "blog_id": "10006274874",
  "title": "夢1歳♡",
  "text": "８月８日❤️\n\n夢空が１歳になりましたぁ🙌💕💕\n\nあの出産から１年…本当に早すぎます💦\n\nだからこそ１日１日を大切にしなくちゃと\n\n改めて思う😌✨✨\n\n夢、毎日沢山の幸せと笑顔を\n\n本当にありがとう👶💕💕✨✨\"\n\nこれからもその笑顔を大切に、\n\n元気でおてんば娘でいてね❤️✨✨\n\n愛してるょ👶💕💕\n\n夢happy birthday🎂💕💕🙌\"\n\n👶💕💕\"\n\n一升餅…可愛かった❤️\n\n賑やかなbirthdayでした🎂💕💕✨\n\n🩷🩷🩷",
  "html": "<div style=\"text-align: center;\"><br></div><div style=\"text-align: center;\"><div>８月８日❤️</div><div>夢空が１歳になりましたぁ🙌💕💕</div><div><p><br></p><div><a href=\"https://stat.ameba.jp/user_images/20260809/00/tsuji-nozomi/28/b4/j/o1080108015810399127.jpg\"><img src=\"https://stat.ameba.jp/use...",
  "published_at": "2026-08-09T00:28:38.000+09:00",
  "updated_at": "2026-08-09T00:28:53.000+09:00",
  "theme_id": "10010811052",
  "theme_name": "ブログ",
  "thumbnail_url": "https://stat.ameba.jp/user_images/20260809/00/tsuji-nozomi/28/b4/j/o1080108015810399127.jpg",
  "tags": [],
  "images": [
    {
      "url": "https://stat.ameba.jp/user_images/20260809/00/tsuji-nozomi/28/b4/j/o1080108015810399127.jpg?caw=800",
      "alt": "",
      "width": 400,
      "height": 400
    },
    {
      "url": "https://stat.ameba.jp/user_images/20260809/00/tsuji-nozomi/2c/fc/j/o1080108015810399131.jpg?caw=800",
      "alt": "",
      "width": 400,
      "height": 400
    },
    {
      "url": "https://stat.ameba.jp/user_images/20260809/00/tsuji-nozomi/2f/33/j/o1080108015810399134.jpg?caw=800",
      "alt": "",
      "width": 400,
      "height": 400
    }
  ],
  "embed_urls": [],
  "is_reblog": false,
  "source_blog_id": null,
  "source_post_id": null,
  "is_member_only": false
}
```

Complete failed result:

```json
{
  "input": "invalid input",
  "status": "failed",
  "id": null,
  "url": null,
  "ameba_id": null,
  "blog_id": null,
  "title": null,
  "text": null,
  "html": null,
  "published_at": null,
  "updated_at": null,
  "theme_id": null,
  "theme_name": null,
  "thumbnail_url": null,
  "tags": null,
  "images": null,
  "embed_urls": null,
  "is_reblog": null,
  "source_blog_id": null,
  "source_post_id": null,
  "is_member_only": null
}
```

A failed row preserves `input`, sets `status` to `failed`, and sets every other top-level field to `null`. The run log records the reason. Optional successful fields can be null or empty when Ameba does not supply them. Object fields are exposed through useful columns in the single **Posts** view; arrays are not expanded into extra result rows.

### Input requirements and pagination settings ⚙️

| Parameter | Type | Required | Default | Rules |
|---|---|---|---|---|
| `inputList` | array of strings | Yes | None | Enter an Ameba blog ID or a public ameblo.jp blog URL. Each page contains up to 20 post records. At least `1` item. |
| `pagesPerBlog` | integer | No | `1` | Maximum result pages per input. Setup requests do not count as pages. Stops when the source has no next page or repeats a continuation. Actual rows per page can vary. Minimum `1`. |

Enter an Ameba blog ID or a public ameblo.jp blog URL. Each page contains up to 20 post records.

Examples of supported inputs: `tsuji-nozomi`.

Duplicate normalized inputs are processed once. Within the collection for one source input, repeated records are removed while distinct interactions or objects remain separate. The same object returned for different source inputs retains its source relationship.

The first result page counts as page 1. Public-page preparation and metadata lookups do not consume result pages. Collection stops at your page budget or the end of the available list. A short page alone does not mean that the list has ended.

### Ameba posts use cases 🎯

#### Build an Ameba lifestyle article catalog

Collect public posts from an Ameba lifestyle creator. Export article text, publication dates, themes and media links to build a source-linked content catalog.

#### Collect recent Ameba service announcements

Read the first page of public Ameba staff blog posts. Keep announcement text, dates and source URLs for a repeatable product update review.

#### Compare content from selected Ameba blogs

Collect public posts from two selected Ameba creator blogs. Compare themes, publication dates and available article text in your own content analysis.

### Pricing and saved-result behavior 💰

Check the Actor’s **Pricing** tab for the active pricing model and current rate. Store settings can change, so this README does not claim a fixed price or runtime.

Under dataset-item pricing:

- Each successful result represents one post for its source input.
- An invalid or unavailable input, an input with no accessible results, or an exhausted request can produce a `failed` row.
- Retry attempts and supporting requests do not create extra dataset rows by themselves.
- Nested media, profile information and other arrays remain part of their parent row.
- A normal empty continuation after earlier successes creates no additional row.
- Saved failed rows are not assumed to be free; check their treatment in the active pricing configuration.

More inputs and larger page budgets can produce more saved rows. Start with a small run and check actual usage before increasing the workload.

### Ameba posts API 🔌

Replace `$ACTOR_ID` with the identifier shown in this Actor’s **API** tab and `$APIFY_TOKEN` with your Apify token.

```bash
curl -X POST "https://api.apify.com/v2/acts/$ACTOR_ID/runs?token=$APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"inputList":["tsuji-nozomi"],"pagesPerBlog":1}'
```

Download the dataset in JSON, CSV, Excel or other formats available in the Console, or retrieve it through the Apify API. Use schedules, webhooks and Apify integrations to connect results to Google Sheets, Make, Zapier, cloud storage or your own application. These connections are configured separately by the user.

### Reliability, retries, and public-data limits ⚠️

Collection uses pure HTTP with up to five concurrent requests. Invalid syntax and confirmed missing, removed or unavailable targets stop without unnecessary retries. Temporary network or proxy failures, timeouts, blocking responses, malformed data, throttling and server errors allow up to five total attempts per request.

Successful earlier pages remain saved if a later request fails. A normal end after saved results is not retried. An input with no accessible results can receive a failed row; that alone does not prove that its underlying source does not exist. An exhausted later request can add a failed row while preserving previous output.

Only publicly returned content is included. Member-only posts can contain public metadata without article text. Deleted or hidden entries can reduce page size.

Ameba controls public availability, ranking and optional fields. Result counts can be affected by duplicates, removed content and changes during a run. A page budget is a collection limit, not a promise of exhaustive coverage.

An invalid string within a valid input list does not stop other inputs. Invalid overall configuration, such as a non-string list item or a wrong page-count type, exits before source requests. Startup failures, unavailable dataset storage or an unrecoverable result-save error can still stop the whole run. A failed save is not retried as a new scraping request.

### Frequently asked questions ❓

#### What does one Ameba Blog Posts Scraper result represent?

Each successful row represents one post. The source input is retained, and nested fields stay in the same row.

#### Which source limits apply?

Only publicly returned content is included. Member-only posts can contain public metadata without article text. Deleted or hidden entries can reduce page size.

#### Does it download media files?

No. Where available, the Actor returns media URLs and metadata within the result row.

#### Does it require a Ameba account or personal cookies?

No. You do not need to provide a Ameba account, password or personal session cookie. Collection uses HTTP without opening a browser.

#### What happens to invalid or unavailable inputs?

Invalid individual inputs are recorded as failed without a source request. Confirmed missing or unavailable targets stop without unnecessary retries. Temporary failures allow up to five total attempts per request; other inputs and earlier saved results remain available.

#### Can I export results or run the same input again?

Yes. Export the default dataset or retrieve it through the Apify API. You can create an Apify schedule and connect completed runs to your own workflow.

### Support, responsible use, and related actors 🛟

For a reproducible issue, use the Actor’s **Issues** tab and provide the run ID, approximate time, a safe public input, expected behavior and actual result. Include the mode or page budget when relevant. Never share access tokens, personal cookies or proxy credentials.

Use public data responsibly and follow applicable privacy, copyright, contractual and platform requirements before storing, analyzing or redistributing collected information.

# Actor input Schema

## `inputList` (type: `array`):

Enter an Ameba blog ID or a public ameblo.jp blog URL. Each page contains up to 20 post records.

## `pagesPerBlog` (type: `integer`):

Maximum result pages per input. Setup requests do not count as pages. Stops when the source has no next page or repeats a continuation. Actual rows per page can vary.

## Actor input object example

```json
{
  "inputList": [
    "tsuji-nozomi"
  ],
  "pagesPerBlog": 1
}
```

# Actor output Schema

## `posts` (type: `string`):

One row per post. Images and tags stay in arrays within that row.. Check success or failed status.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "inputList": [
        "tsuji-nozomi"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("scrapingmonkey/ameba-blog-posts-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "inputList": ["tsuji-nozomi"] }

# Run the Actor and wait for it to finish
run = client.actor("scrapingmonkey/ameba-blog-posts-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "inputList": [
    "tsuji-nozomi"
  ]
}' |
apify call scrapingmonkey/ameba-blog-posts-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,scrapingmonkey/ameba-blog-posts-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/ag0iuNkH9RxmF1bjQ/builds/KGIy9ocAfEHeqcbsH/openapi.json
