# Google News Scraper (`maged120/google-news-scraper`) Actor

Scrape Google News headlines by keyword, topic or link in 28 country editions, with source, publish time and the original article URL.

- **URL**: https://apify.com/maged120/google-news-scraper.md
- **Developed by:** [Maged](https://apify.com/maged120) (community)
- **Categories:** News, Marketing, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

**Scrape Google News by keyword, topic or link.** Get every headline with its **source**, **publish time** and the **original article URL** (the real publisher link, not a Google redirect). Covers 28 country editions and works great on a schedule for news monitoring.

### What does Google News Scraper do?

Google News Scraper collects articles from [Google News](https://news.google.com/) for any keyword, topic or Google News page you choose. Each article becomes a **clean row**: headline, publisher, publisher website, publish time and a direct link to the original story.

It runs on the Apify platform, so you also get API access, scheduled runs, integrations (Google Sheets, Slack, Zapier, Make, webhooks) and run monitoring.

### Why use Google News Scraper?

- **Brand and competitor monitoring**: track every mention of your company, product or competitors.
- **Media and PR tracking**: see which publications cover your space and how often.
- **Market and investment research**: follow news on companies, industries and keywords in real time.
- **Content and SEO**: spot trending stories in your niche for fast, relevant content.
- **Data and AI pipelines**: feed fresh, structured news into dashboards, alerts or language models.

### How to scrape Google News

1. Open the **Input** tab.
2. Add **keywords** (`electric vehicles`), **topics** (`TECHNOLOGY`) or **Google News links**.
3. Choose an **edition** (country and language) and a **time range**.
4. Click **Start**. Articles appear in the **Output** tab within seconds.
5. Download the results as JSON, CSV, Excel or HTML, or connect them to Slack, Sheets or your app.

### Input

| Field | Description |
|---|---|
| `searches` | Keywords, topics (WORLD, NATION, BUSINESS, TECHNOLOGY, ENTERTAINMENT, SPORTS, SCIENCE, HEALTH) or Google News links. Empty = top stories. |
| `edition` | Country and language edition. Default: United States (English). |
| `timeRange` | Any time, past hour, 24 hours, 7 days, 30 days or year. |
| `maxArticlesPerSearch` | Up to 100 articles per search. |
| `resolveArticleUrls` | Return the original publisher URL. Default on. |

```json
{
    "searches": ["\"electric vehicles\" site:reuters.com", "TECHNOLOGY", "OpenAI OR Anthropic"],
    "edition": "US",
    "timeRange": "1d",
    "maxArticlesPerSearch": 100
}
```

Keyword searches support Google News operators: quotes for exact phrases, `OR`, `-word` to exclude, `site:` to limit to one publisher and `intitle:` for headline matches.

### Output

Each article is one row. You can download the dataset in various formats such as JSON, HTML, CSV, or Excel.

```json
{
    "entityType": "article",
    "searchTerm": "electric vehicles",
    "position": 1,
    "title": "Batteries, Charging, and Electric Vehicles",
    "source": "Department of Energy (.gov)",
    "sourceUrl": "https://www.energy.gov",
    "publishedAt": "2026-09-24T14:21:45+00:00",
    "articleUrl": "https://www.energy.gov/cmei/vehicles/batteries-charging-and-electric-vehicles",
    "googleNewsUrl": "https://news.google.com/rss/articles/CBMigwFBVV95cUxQM0ZZMG5..."
}
```

#### Data fields

| Field | What it means |
|---|---|
| `title` | Headline (without the trailing publisher name) |
| `source`, `sourceUrl` | Publisher name and website |
| `publishedAt` | Publish time (ISO 8601, UTC) |
| `articleUrl` | Original article link on the publisher's site |
| `googleNewsUrl` | The Google News link for the same story |
| `position` | Rank in the Google News results |
| `searchTerm` | The keyword, topic or link that found it |

### How many results will I get?

Up to 100 articles per keyword, topic or link. For more coverage, run several narrower searches, for example the same keyword with different `site:` filters or editions, or schedule frequent runs with a short time range.

### Tips

- **Alerts**: schedule hourly with `timeRange: 1h` and connect Slack or email to get new mentions as they happen.
- **Faster runs**: turn off `resolveArticleUrls` if Google News links are enough for you.
- **Exact matches**: put brand names in quotes (`"Acme Corp"`) to avoid unrelated results.
- **Local news**: pick the edition of the country you care about. Results differ a lot between editions.

### FAQ

**Is it legal to scrape Google News?**
The actor collects publicly available headlines, sources, dates and links. It doesn't copy full article text. You are responsible for how you use the data and for respecting publishers' copyrights.

**Does it return the full article text?**
No. It returns headlines, sources, times and links. Pair it with an article extractor if you need the full text.

**Why are only 100 articles returned?**
Google News lists at most 100 articles per search. Use narrower or multiple searches for more.

**Something missing or broken?**
Open an issue in the **Issues** tab and we'll look at it quickly. Need a custom news-monitoring setup? Get in touch through the same tab.

***

⭐ **Found this Actor useful?** A quick review on the Store helps other users find it and keeps it maintained.

# Actor input Schema

## `searches` (type: `array`):

One per line. A keyword or phrase ("electric vehicles"; search operators like site:, intitle: and OR work), a topic (WORLD, NATION, BUSINESS, TECHNOLOGY, ENTERTAINMENT, SPORTS, SCIENCE, HEALTH), or a Google News link from your browser. Leave empty for today's top stories.

## `edition` (type: `string`):

Which Google News edition to read.

## `timeRange` (type: `string`):

Only articles published within this period (keyword searches).

## `maxArticlesPerSearch` (type: `integer`):

Up to 100 articles per keyword, topic or link.

## `resolveArticleUrls` (type: `boolean`):

Return the publisher's real article URL instead of a Google News redirect link. Turn off for faster runs.

## Actor input object example

```json
{
  "searches": [
    "electric vehicles",
    "TECHNOLOGY"
  ],
  "edition": "US",
  "timeRange": "any",
  "maxArticlesPerSearch": 100,
  "resolveArticleUrls": true
}
```

# Actor output Schema

## `dataset` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "searches": [
        "electric vehicles",
        "TECHNOLOGY"
    ],
    "maxArticlesPerSearch": 100
};

// Run the Actor and wait for it to finish
const run = await client.actor("maged120/google-news-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "searches": [
        "electric vehicles",
        "TECHNOLOGY",
    ],
    "maxArticlesPerSearch": 100,
}

# Run the Actor and wait for it to finish
run = client.actor("maged120/google-news-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "searches": [
    "electric vehicles",
    "TECHNOLOGY"
  ],
  "maxArticlesPerSearch": 100
}' |
apify call maged120/google-news-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,maged120/google-news-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/cECw5zExeJpWJPJde/builds/eMpMjWl0l8Xa1Ib3M/openapi.json
