# Hotel Reviews Scraper (Google, Booking, Tripadvisor) (`scrapercompany/hotel-reviews-scraper`) Actor

Scrape hotel guest reviews from Google Hotels, Booking.com, Expedia, Hotels.com and Tripadvisor in one Actor: rating, title, text, date, stay date, room type, trip type, reviewer country and the hotel's reply. Paging, sorting and keyword filter. No proxies needed.

- **URL**: https://apify.com/scrapercompany/hotel-reviews-scraper.md
- **Developed by:** [ScraperCompany](https://apify.com/scrapercompany) (community)
- **Categories:** Travel, SEO tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $3.00 / 1,000 reviews pages

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Hotel Reviews Scraper (Google, Booking, Tripadvisor)

Hotel Reviews Scraper collects guest reviews for hotels from **five sources** with one consistent output: Google Hotels (which aggregates many providers), Booking.com, Expedia, Hotels.com and Tripadvisor. Every review keeps its rating and the source's rating scale, title and text, review date and stay date, room type, trip type, reviewer country, helpful votes and the hotel's reply. Use it for reputation monitoring, sentiment analysis, competitor benchmarking and training data.

Powered by the [ScraperCompany API](https://scrapercompany.com): requests, proxies, retries and anti-bot handling run on ScraperCompany's servers, so you don't need your own proxies or API key. Start a run, get structured JSON.

### What can this scraper do?

- **Five sources, one schema**: Google Hotels, Booking.com, Expedia, Hotels.com and Tripadvisor.
- **Full review data**: rating (with the source's scale), title, text, language, review and stay dates, room type, trip type, reviewer country, helpful votes.
- **Management replies** included when the hotel answered.
- **Paging**: up to 50 requests per hotel; Google returns up to 100 reviews per request.
- Booking.com **sort** (newest, oldest, relevant) and **keyword filter**.
- Empty pages are never charged.

#### Sources

| Source | What it returns | Price per request |
| --- | --- | --- |
| `google` | Google Hotels reviews from every provider Google aggregates (up to 100 per request) | $0.003 |
| `booking` | Booking.com reviews with 0-10 scores, room type, stay dates and the hotel's reply | $0.003 |
| `expedia` | Expedia verified-guest reviews | $0.003 |
| `hotels` | Hotels.com verified-guest reviews | $0.003 |
| `tripadvisor` | Tripadvisor reviews with 1-5 bubbles, trip type and the management response | $0.003 |

### What data can you extract?

| Field | Description |
| --- | --- |
| `reviews[].rating / rating_scale` | Score and the scale it is on (10 for Booking.com/Expedia/Hotels.com, 5 for Tripadvisor) |
| `reviews[].title / text / language` | Review content |
| `reviews[].date / stay_date` | When it was written and when the guest stayed |
| `reviews[].room_type / trip_type` | Room booked and trip type (family, couple, business...) |
| `reviews[].author / author_country` | Reviewer display name and country, as published |
| `reviews[].management_reply` | The hotel's public answer |
| `reviews[].provider / source / url` | Where the review comes from |
| `total / count / next_page_token` | Totals and paging cursor |
| `rating_scores` | Booking.com category scores (cleanliness, location, staff...) |

Every dataset item also carries `input` (the exact request that was sent), `request_id` (quote it to support), `billed` and `scraped_at`.

### How to use Hotel Reviews Scraper (Google, Booking, Tripadvisor)

1. Open Hotel Reviews Scraper (Google, Booking, Tripadvisor) in Apify Console and go to the **Input** tab.
2. Pick a **Review source**, then enter hotels the way that source identifies them (Google property token, Booking.com slug, Expedia/Hotels.com id or Tripadvisor location id). Set **Pages per hotel** for more reviews.
3. Adjust the options if needed. Dates accept an exact date (`2026-12-01`) or a relative one (`30 days` from today), so saved tasks and schedules never go stale.
4. Click **Start** and wait for the run to finish.
5. Download the results from the **Output** tab as JSON, CSV, Excel or HTML, or fetch them with the Apify API.

### Input example

This is the default input; running the Actor without changes uses it.

```json
{
  "source": "google",
  "properties": [
    "ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"
  ],
  "maxPages": 1,
  "reviewsPerPage": 10,
  "includeErrors": true,
  "maxConcurrency": 3,
  "maxRetries": 2
}
```

### Output example

One dataset item per request (trimmed here; real items contain every field the API returns):

```json
{
  "source": "booking",
  "endpoint": "/v1/ota/booking/reviews",
  "input": {
    "pagename": "newbury-guest-house",
    "country": "us",
    "limit": 10,
    "sort": "MOST_RELEVANT"
  },
  "collection_id": "ratecol_0123456789abcdef0123456789abcdef",
  "count": 2,
  "country": "us",
  "elapsed_s": 0.521,
  "hotel_id": 236725,
  "limit": 10,
  "observed_at": "2026-08-07T01:23:45.678Z",
  "pagename": "newbury-guest-house",
  "rating_scale": 10,
  "rating_scores": [
    {
      "count": 258,
      "name": "Staff",
      "translation": null,
      "value": 9.3
    }
  ],
  "reviews": [
    {
      "author": "Anna",
      "author_country": "United States",
      "date": "2024-05-23",
      "helpful_votes": 0,
      "language": "xu",
      "management_reply": "Hello Anna, we are pleased that you enjoyed your stay. We value your kind words and feedback.",
      "provider": null,
      "rating": 8,
      "review_id": "48b4dc6c098c149e",
      "room_type": "Deluxe Room",
      "source": "booking",
      "stay_date": "2024-05-23",
      "text": "No breakfast but coffee/tea in lobby, location very good-on fanciest street in Boston and room was clean neat and comfortable.",
      "title": "Good value, warm and inviting staff and very neat and clean rooms.",
      "trip_type": null,
      "url": "48b4dc6c098c149e"
    }
  ],
  "skip": 0,
  "sort": "MOST_RELEVANT",
  "sorters": [
    {
      "name": "Most relevant",
      "value": "MOST_RELEVANT"
    }
  ],
  "total": 523,
  "ufi": 20061717,
  "wire_bytes": 14973,
  "request_id": "req_3f9c0d6e2b8a4c1f9e7d5b3a1c0e8f6d",
  "billed": true,
  "scraped_at": "2026-10-01T14:03:27.512Z"
}
```

A request that fails is saved with an `error` message instead (turn off **Include failed requests** to skip those). Failed requests are never charged.

### How much does it cost?

This Actor uses **pay-per-event** pricing: you pay for successful API requests, not for compute time.

| Event | Charged when | Price | Per 1,000 |
| --- | --- | --- | --- |
| `review-page` | One page of guest reviews from one source (Google: up to 10 pages of 10 per request). Empty pages are free. | $0.003 | $3.00 |

For example, 1,000 review pages cost **$3.00**. One Google request can return up to 100 reviews (10 pages) for the price of one page.

- Failed requests (errors, invalid input, blocked upstream after retries) are **free**.
- Requests where the source returns nothing at all (no results, nothing priced) are saved but **not charged**.
- Set **Maximum cost per run** when you start a run and the Actor stops cleanly before going over it.

### FAQ

#### Do I need proxies or a ScraperCompany API key?

No. Proxy rotation, retries and anti-bot handling run on the ScraperCompany side, and the Actor is already connected to the API. You only pay the per-event prices above.

#### How do I find the hotel identifiers?

Google property tokens come from the Google Hotels Search Scraper; Booking.com slugs from the Booking.com URL; Expedia and Hotels.com ids from their URLs or their scrapers' Find property id mode; Tripadvisor location ids from the URL (`-d208453-`) or the Tripadvisor scraper.

#### Are reviewer names personal data?

They are published publicly by the source, but they can still be personal data under GDPR. Only collect and store them if you have a lawful basis.

#### Is it legal to scrape this data?

The Actor collects publicly available information that anyone can see without logging in. You are responsible for how you use the results: respect the source site's terms, copyright and privacy law (such as GDPR) and do not collect personal data without a lawful basis. If in doubt, ask a lawyer.

#### What happens when a request is blocked or rate-limited?

Rate limits (HTTP 429) and temporary errors (5xx) are retried automatically with exponential backoff, respecting `Retry-After`. If a request still fails it is saved with the error message and not charged, and the rest of the batch keeps going. A run only fails when every request failed.

#### Why did a request return an error?

Read the `error` field: validation problems (for example a malformed date or an unknown id) are reported exactly as the API sees them. Fix the input and run again; failed requests cost nothing.

#### How many requests can I run at once?

Any number per run. The Actor sends up to 3 requests in parallel by default (change it under **Run options**); larger batches simply take longer.

### Use it from your code

Call the Actor from any language through the [Apify API](https://docs.apify.com/api/v2). With the JavaScript client (`npm install apify-client`):

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('scrapercompany/hotel-reviews-scraper').call({
    "source": "google",
    "properties": [
        "ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"
    ]
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);
```

With Python (`pip install apify-client`):

```python
import os
from apify_client import ApifyClient

client = ApifyClient(os.environ["APIFY_TOKEN"])
run = client.actor("scrapercompany/hotel-reviews-scraper").call(run_input={
    "source": "google",
    "properties": ["ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"],
})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)
```

You can also schedule runs, chain them with webhooks, or connect them to Make, Zapier, n8n, Google Sheets and other integrations from the **Integrations** tab.

### Related scrapers

- [Google Hotels Search Scraper](https://apify.com/scrapercompany/google-hotels-search-scraper)
- [Tripadvisor Hotel Prices Scraper](https://apify.com/scrapercompany/tripadvisor-hotel-rates-scraper)
- [Booking.com Hotel Rates Scraper](https://apify.com/scrapercompany/booking-hotel-rates-scraper)

### Support

Questions, a field you need, or a site that stopped working? Open an issue on the **Issues** tab or contact us at [scrapercompany.com](https://scrapercompany.com). Include the `request_id` from the dataset item so we can trace the request.

# Changelog

This Actor's version history is a separate document: https://apify.com/scrapercompany/hotel-reviews-scraper/changelog.md

# Actor input Schema

## `source` (type: `string`):

Where to read reviews from. Each source identifies hotels differently (see Hotels).

## `properties` (type: `array`):

One hotel per line, identified the way the selected source does: <b>Google</b> property token (<code>ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ</code>), <b>Booking.com</b> slug (<code>newbury-guest-house</code>), <b>Expedia</b> / <b>Hotels.com</b> numeric id (<code>12570</code>), <b>Tripadvisor</b> location id (<code>208453</code>).

## `maxPages` (type: `integer`):

Requests per hotel (1-50). Each request is one page and is charged separately; paging stops early at the last page. For Google, also see <i>Google pages per request</i>.

## `reviewsPerPage` (type: `integer`):

Reviews per request: up to 25 for Booking.com, Expedia and Hotels.com, up to 30 for Tripadvisor.

## `pages` (type: `integer`):

Google only: how many 10-review pages to follow in one request (1-10, up to 100 reviews for one charge).

## `gl` (type: `string`):

Two-letter Google country for the request.

## `hl` (type: `string`):

Language for the request, e.g. <code>en</code>, <code>fr</code>.

## `country` (type: `string`):

Two-letter segment before the slug in the Booking.com URL (<code>us</code> in <code>/hotel/us/newbury-guest-house.html</code>).

## `sort` (type: `string`):

Order of Booking.com reviews.

## `text` (type: `string`):

Only reviews mentioning this keyword (filtered by Booking.com).

## `expedia_market` (type: `string`):

Expedia point of sale (US or CA).

## `hotels_market` (type: `string`):

Hotels.com point of sale.

## `customRequests` (type: `array`):

Optional list of JSON objects, one request each, using the API field names shown above. Each object is merged over the options above, so you only need to give what differs. Use <code>"source"</code> (google, booking, expedia, hotels, tripadvisor) to pick the source per object, with that source's id field (<code>property_token</code>, <code>pagename</code>, <code>property_id</code> or <code>location_id</code>). Example: <code>{"source":"booking","pagename":"newbury-guest-house","country":"us","sort":"NEWEST_FIRST"}</code>

## `includeErrors` (type: `boolean`):

When on, a request that fails (for example an unknown property id) is saved as a dataset item with an <code>error</code> message so you can see what went wrong. Failed requests are never charged.

## `maxConcurrency` (type: `integer`):

How many API requests run at the same time.

## `maxRetries` (type: `integer`):

Retries for rate limits (HTTP 429), temporary server errors (5xx) and network errors, with exponential backoff. Validation errors are never retried.

## Actor input object example

```json
{
  "source": "google",
  "properties": [
    "ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"
  ],
  "maxPages": 1,
  "reviewsPerPage": 10,
  "pages": 1,
  "gl": "us",
  "hl": "en",
  "country": "us",
  "sort": "MOST_RELEVANT",
  "expedia_market": "US",
  "hotels_market": "US",
  "includeErrors": true,
  "maxConcurrency": 3,
  "maxRetries": 2
}
```

# Actor output Schema

## `results` (type: `string`):

One dataset item per request.

## `summary` (type: `string`):

Counts of succeeded, failed, skipped and charged requests.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "properties": [
        "ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("scrapercompany/hotel-reviews-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "properties": ["ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"] }

# Run the Actor and wait for it to finish
run = client.actor("scrapercompany/hotel-reviews-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "properties": [
    "ChUIoben2Mv6-CYaCi9tLzA3czVwbjQQAQ"
  ]
}' |
apify call scrapercompany/hotel-reviews-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,scrapercompany/hotel-reviews-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/WqreAHrh1AiVN0EJJ/builds/YonhvKDJTtjxiD5vG/openapi.json
