# Google News Scraper (`intelscrape/google-news-scraper`) Actor

Extract Google News headlines, publishers, publish times, snippets, and links via public RSS queries. Pay per article. Demo mode included.

- **URL**: https://apify.com/intelscrape/google-news-scraper.md
- **Developed by:** [IntelScrape](https://apify.com/intelscrape) (community)
- **Categories:** AI, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.00 / 1,000 results

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Google News Scraper – Headlines Publishers & Article Links

> 🇨🇳 **中文简介 (Chinese — China & Singapore):**
> 需要按关键词监控 Google News 头条？本 Actor 通过公开 RSS 提取标题、媒体、发布时间、摘要与链接，按交付文章计费，无需 API Key——适合品牌监测、舆情与 AI 资讯管道。立刻在 Apify 试跑，把新闻流接入 n8n 或 Claude。

> 🇸🇬 **Ringkasan Melayu (Singapore):**
> Perlu pantau tajuk Google News ikut kata kunci? Actor ini ekstrak tajuk, penerbit, masa, snippet dan pautan melalui RSS awam. Bayar per artikel dihantar, tiada API key. Sesuai untuk pemantauan jenama dan penyelidikan di Singapura. Cuba di Apify Store sekarang.

#### 🤖 NEW: Connect This To Your AI!

Want to use this scraper directly inside Claude, Cursor, or ChatGPT? Check out **[Skip Trace MCP Server](https://apify.com/intelscrape/skip--trace)**. Combined IntelScrape agents over Model Context Protocol — autonomous leads, enrichment, outreach prep.

> **Soft CTA:** Need company review intelligence next? Pair with **[Trustpilot Reviews Scraper](https://apify.com/intelscrape/trustpilot-reviews-scraper)** (`review-summary` PPE) + **[Skip Trace PRO](https://apify.com/intelscrape/skip-trace-pro)** for people behind the story.

#### SoftCTA (dataset)

Each `query-summary` row includes a `softCta` object with next-step Actor links (Contact Info, Website Leads, Trustpilot, Skip Trace PRO) — funnel mesh, not a separate charge.

#### 🏆 Featured Bots by IntelScrape

1. **[Google News Scraper](https://apify.com/intelscrape/google-news-scraper)** — Headlines, Publishers & Article Links
2. **[Skip Trace PRO](https://apify.com/intelscrape/skip-trace-pro)** — Name, Address, Phone & Email People Lookup
3. **[TruePeopleSearch Scraper](https://apify.com/intelscrape/truepeoplesearch-scraper)** — Phone & Email Matches
4. **[TikTok Scraper](https://apify.com/intelscrape/tiktok-scraper)** — Emails & Influencer Leads
5. **[Amazon Product Review Scraper](https://apify.com/intelscrape/amazon-product-review-scraper)** — Deep E-commerce & Review Extraction
6. **[Building Permit Scraper](https://apify.com/intelscrape/building-permit-scraper)** — US Cities Open Data
7. **[Google Maps No-Website Leads](https://apify.com/intelscrape/website-Leads)** — Local Businesses Without a Website
8. **[Skip Trace MCP Server](https://apify.com/intelscrape/skip--trace)** — Use IntelScrape inside Claude, Cursor, ChatGPT

***

⚡ **Use this Actor in n8n — no code**

1. Add the official **Apify** node in n8n.
2. Connect your Apify API token.
3. Run Actor ID: `IntelScrape/google-news-scraper`.

***

![Google News Scraper Hero Banner](https://api.apify.com/v2/key-value-stores/PyvJzfQg02FD3RYhd/records/google-news-scraper-banner.png)

![price](https://img.shields.io/badge/price-$1.00_/_1K-e83e8c)
![billing](https://img.shields.io/badge/billing-pay_per_article-2ea44f)
![API key](https://img.shields.io/badge/API_key-not_required-blue)
![coverage](https://img.shields.io/badge/coverage-Google_News_RSS-8a2be2)
![demo](https://img.shields.io/badge/demoMode-fixtures_included-informational)

**Google News Scraper** — clean public-page extraction with honest pay-per-result billing.\
**≈ $$1.00/1K + $0.005 actor start. No API key required.**

> **Honest coverage & billing (read this):** Source is **public Google News / RSS** for your queries — not full-text paywalled article bodies. Billed **$0.005** start + **$0.001** per delivered article + optional **$0.01** `news-summary` per query when `includeNewsSummary` is on (default). Summaries are heuristic (publishers / keywords / date range) — **no LLM**. `demoMode` = fictional fixtures.

#### Pricing (PPE)

| Event | Price | When charged |
| :--- | :--- | :--- |
| Actor start | $0.005 | Once per run |
| Article | $0.001 | Each delivered article row |
| News query summary | $0.01 | Each query-summary row (`includeNewsSummary`) |

***

> 🧭 Built for **brand monitoring, researchers, journalists, and AI news pipelines.** Popular with teams in the **US, China, and Singapore**.

> ⭐ **Store rating is driven by daily power users.** If this Actor saves you time, please [leave a review](https://apify.com/intelscrape/google-news-scraper).

> ⚖️ **Lawful business use only.** Public pages only. You are liable for how you use the data.

> 📌 *Examples in this README are **fictional** unless stated otherwise.*

***

### 🕵️‍♂️ Why Choose Us vs. "Cheaper" Competitors?

| Feature | Google News Scraper (IntelScrape) | Thin scrapers |
| :--- | :--- | :--- |
| **Source** | Public News RSS | Brittle SERP HTML |
| **Billing** | Pay per delivered article | Empty row charges |
| **Automation** | MCP + n8n documented | Manual copy |
| **Region / time / topics** | `regionLanguage` + `timeframe` + topic feeds | US-only thin RSS |

***

### Filters (rival parity)

| Input | What it does |
| :--- | :--- |
| `topics[]` | WORLD / NATION / BUSINESS / TECHNOLOGY / ENTERTAINMENT / SPORTS / SCIENCE / HEALTH topic feeds |
| `timeframe` | `6h` / `1h` / `1d` / `7d` / `30d` / `1y` / `all` (RSS `when:` + client filter) |
| `dateFrom` / `dateTo` | YYYY-MM-DD calendar window; overrides timeframe for client filter when set (crawlerbros/easyapi parity) |
| `maxResultsPerQuery` | Optional per-query/topic cap before global `maxArticles` |
| `regionLanguage` | e.g. `US:en`, `GB:en`, `DE:de` → hl/gl/ceid |
| `siteFilter[]` | Append `site:domain` to keyword queries |
| `topicUrls[]` | Custom Google News topic/section URLs → RSS (data\_xplorer parity) |
| `extractDescriptions` / `extractImages` | Optional page enrich for meta description / og:image |
| `filtersApplied` | Echo of active filters on every row |
| `excludeWords[]` | Append `-word` exclusions |
| `decodeUrls` | Normalize article URLs (default on) |
| `includeNewsSummary` | Extractive query-summary row + `news-summary` PPE |

### Quick start

```json
{
  "queries": ["artificial intelligence"],
  "maxArticles": 50,
  "includeNewsSummary": true,
  "demoMode": false
}
```

**Demo (fictional fixtures):**

```json
{ "demoMode": true }
```

***

### What you get

- `articleId`
- `title`
- `publisher`
- `publishedAt`
- `snippet`
- `url`
- `query`
- `source`

***

### Pricing (PPE — live)

| Event | Price | When it fires |
| :--- | ---: | :--- |
| Actor start (`apify-actor-start`) | **$0.005** | Once per run |
| Article (`apify-default-dataset-item`) | **$0.001** | Each delivered row via `pushData` |

***

### Legal

Public pages only. Not affiliated with the source sites. Lawful B2B / research use; you are responsible for compliance. Demo content is fictional.

# Actor input Schema

## `queries` (type: `array`):

Google News search queries (e.g. electric vehicles). Supports operators: OR, quotes, site:, -exclude.

## `topics` (type: `array`):

Predefined Google News topic feeds (in addition to keyword queries).

## `topicUrls` (type: `array`):

Custom Google News topic or section URLs (data\_xplorer parity). Examples: https://news.google.com/topics/CAAq… or /rss/topics/…. Scraped via RSS. Cap 20.

## `timeframe` (type: `string`):

Recency filter for keyword searches (appended as when: to the RSS query + client filter). Topics ignore RSS when:; client filter still applies unless dateFrom/dateTo override. Use dateFrom/dateTo for exact calendar bounds.

## `dateFrom` (type: `string`):

Only keep articles published on/after this date (YYYY-MM-DD). Overrides/narrows timeframe for client filtering when set. Rival parity with crawlerbros/powerai/easyapi.

## `dateTo` (type: `string`):

Only keep articles published on/before this date (YYYY-MM-DD). Use with dateFrom for a calendar window.

## `maxResultsPerQuery` (type: `integer`):

Optional per-query/topic cap applied before the global maxArticles limit (crawlerbros parity). Leave unset for no per-query cap.

## `regionLanguage` (type: `string`):

Google News region:language (gl/hl/ceid). Default US English.

## `siteFilter` (type: `array`):

Limit keyword searches to these domains (appended as site:domain). Rival parity with crawlerbros/memo23.

## `excludeWords` (type: `array`):

Words to exclude from keyword searches (appended as -word).

## `decodeUrls` (type: `boolean`):

Normalize article URLs via URL parser (best-effort; full CBM decode not billed separately).

## `maxArticles` (type: `integer`):

Maximum articles to push across all queries/topics.

## `extractDescriptions` (type: `boolean`):

Fetch each article page and pull meta description / og:description when snippet is thin (data\_xplorer parity). Slower.

## `extractImages` (type: `boolean`):

Fetch og:image from article pages when RSS has no image (data\_xplorer/crawlerbros parity). Slower.

## `maxEnrich` (type: `integer`):

Cap how many article pages get extractDescriptions/extractImages fetches (protects run time). Default 20.

## `demoMode` (type: `boolean`):

Push synthetic fixtures (no live scrape).

## `includeNewsSummary` (type: `boolean`):

Push one extractive query-summary row per search query/topic (top publishers, keywords, date range) and charge news-summary. Default on.

## Actor input object example

```json
{
  "queries": [
    "artificial intelligence"
  ],
  "topics": [],
  "topicUrls": [],
  "timeframe": "all",
  "regionLanguage": "US:en",
  "siteFilter": [],
  "excludeWords": [],
  "decodeUrls": true,
  "maxArticles": 50,
  "extractDescriptions": false,
  "extractImages": false,
  "maxEnrich": 20,
  "demoMode": false,
  "includeNewsSummary": true
}
```

# Actor output Schema

## `articles` (type: `string`):

Article rows pushed to the default dataset.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "queries": [
        "artificial intelligence"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("intelscrape/google-news-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "queries": ["artificial intelligence"] }

# Run the Actor and wait for it to finish
run = client.actor("intelscrape/google-news-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "queries": [
    "artificial intelligence"
  ]
}' |
apify call intelscrape/google-news-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,intelscrape/google-news-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/52ZgnSjjMS6VMAow1/builds/V8t3r72eECAB4ubiR/openapi.json
