Google News Scraper — Headlines, Sources, URLs (2K Max) avatar

Google News Scraper — Headlines, Sources, URLs (2K Max)

Pricing

from $5.00 / 1,000 results

Go to Apify Store
Google News Scraper — Headlines, Sources, URLs (2K Max)

Google News Scraper — Headlines, Sources, URLs (2K Max)

Turn any Google News query into a deduplicated dataset of up to 2,000 articles: headlines, sources, dates, RSS links, publisher URLs, and snippets. No API key. Apify AI, MCP, and Cursor ready.

Pricing

from $5.00 / 1,000 results

Rating

5.0

(4)

Developer

Shop Intel

Shop Intel

Maintained by Community

Actor stats

3

Bookmarked

52

Total users

11

Monthly active users

21 hours ago

Last modified

Categories

Share

Google News Scraper – Extract Headlines, Sources & Links by Keyword | Apify Actor

Google News Scraper collects headlines, publication dates, sources, snippets, and direct article URLs from Google News for any keyword — export up to 2,000 deduplicated results per run to CSV or JSON, Excel‑ready out of the box.

Use with Apify AI

Apify AI is live in Apify Console. Describe the job in plain English and it finds this Actor, fills the input form, runs it, and returns the dataset. The same ranking powers Store search and the Apify MCP server search-actors tool used by Cursor, Claude, ChatGPT, and other agents.

Use this Actor to turn a Google News query into headlines, sources, dates, and article URLs. Open Google News Scraper or ask Apify AI with the prompts below.

Prompts that match this Actor

Type these in the Apify Store search bar (long, intent-heavy queries route to Apify AI) or the dashboard Ask Apify AI widget:

  • "Collect Google News headlines about Apple earnings"
  • "Scrape the latest news articles for climate change"
  • "Monitor Google News for a brand keyword and export URLs"
  • "Get 200 Google News results for electric vehicles"
  • "Build a news dataset of headlines and publisher links for a topic"

How Apify AI fills the input

You sayThis Actor sets
topic or brandkeyword
how many unique articles (max 2000)numberOfResults

Apify AI always asks for confirmation before it runs. Nothing is charged until you approve.

Cursor, Claude, ChatGPT (Apify MCP)

Pin this Actor as a default Cursor / MCP tool so agents call scrapeio/google-news-scraper instead of a random Store result. Keep actors + docs so Apify AI search still works, and list your suite so Cursor prefers these Actors.

Cursor MCP URL (this Actor first):

https://mcp.apify.com?tools=actors,docs,scrapeio/google-news-scraper,scrapeio/amazon-scraper,scrapeio/google-maps-scraper-advance,scrapeio/meta-facebook-ad-scrapper-using-ad-library-url-premium,scrapeio/instagram-scraper-premium,scrapeio/whatsapp-scraper-premium,scrapeio/facebook-ad-library-suggestions

.cursor/mcp.json / Cursor Settings → MCP:

{
"mcpServers": {
"apify": {
"url": "https://mcp.apify.com?tools=actors,docs,scrapeio/google-news-scraper,scrapeio/amazon-scraper,scrapeio/google-maps-scraper-advance,scrapeio/meta-facebook-ad-scrapper-using-ad-library-url-premium,scrapeio/instagram-scraper-premium,scrapeio/whatsapp-scraper-premium,scrapeio/facebook-ad-library-suggestions"
}
}
}

Then: search-actorsfetch-actor-detailscall-actor with scrapeio/google-news-scraper. Example:

Use scrapeio/google-news-scraper to do this: scrape the data I described. Fill the input from my description and return the dataset.

How to scrape Google News

Use this Google News scraper to extract headlines, sources, and article URLs by keyword. Up to 2,000 unique articles. No Google API key.

Google News scraper API for Cursor

Call scrapeio/google-news-scraper from Cursor (Apify MCP), Claude, ChatGPT, the Apify API, Python (apify-client), or JavaScript. Apify AI uses the same ranking as Store search: title, description, README, and input/output schemas. This Actor is documented for that ranker — limited permissions, pay-per-event or compute (not rental), and a filled example input.

🧠 Overview

This Google News Scraper is an Apify Actor built for journalists, PR agencies, SEO analysts, investors, and data teams who need real‑time news monitoring at scale without the cost or complexity of Google's paid News API. Point it at a keyword, set how many unique articles you need (1–2,000), and the Actor pulls from the public Google News RSS layer, merges multiple time windows (1h, 7d, 30d, 1y, all), follows pagination automatically, and deduplicates by GUID and link. Perfect for brand monitoring, competitive intelligence, PR measurement, stock market sentiment tracking, and research datasets — no API keys, no quotas, no rate‑limited SDKs.

✨ Features

  • Extract headlines, publish dates, source names, snippets, and direct article URLs in one structured Dataset.
  • Search any keyword or phrase just like typing into Google News.
  • Scrape up to 2,000 unique articles per run with automatic deduplication across feed passes.
  • Merge multiple RSS time windows (1h, 24h, 7d, 30d, 1y, all) to maximize depth on a single keyword.
  • Unwrap Google's redirect URLs to return the best‑effort publisher URL when available.
  • Export Excel‑friendly CSV (UTF‑8 BOM, quoted fields, CRLF) plus JSON in one run.
  • Monitor run diagnostics via OUTPUT.meta.stoppedReason (completed, partial_no_more_feed, no_items_or_http_error).
  • Automate with the Apify scheduler for daily, hourly, or custom‑cron news pipelines.
  • Integrate with n8n, Make, Zapier, Google Sheets, and BigQuery via the Apify REST API and webhooks.
  • Skip Google Cloud setup — no News API key, no billing, no OAuth.

🎯 Use Cases

  • Brand Monitoring: Track every mention of your company, product, or executives across thousands of publishers in real time.
  • Competitive Intelligence: Watch competitor product launches, hiring announcements, and funding events by scraping their name or ticker daily.
  • PR & Media Measurement: Measure share of voice, earned‑media coverage, and campaign pickup across trade and mainstream press.
  • Investment & Stock Research: Feed headline datasets into sentiment models for tickers, commodities, or macro events like Fed rate decision.
  • SEO & Content Planning: Spot trending topics, breaking news, and publisher angles for timely SEO content and editorial calendars.
  • AI Training Datasets: Build headline or news‑summary corpora for LLM fine‑tuning, classification, or retrieval‑augmented generation.

⚙️ Input Parameters

NameTypeRequiredDescriptionExample
keywordstringYesSearch term or phrase (aliases: q, query, searchQuery)."artificial intelligence"
numberOfResultsintegerYesUnique articles to collect, 1–2000 (aliases: maxResults, results).200

📤 Output Example (JSON)

{
"position": 1,
"keyword": "artificial intelligence",
"title": "OpenAI unveils next‑gen reasoning model for enterprise customers",
"link": "https://news.google.com/rss/articles/CBM...",
"articleUrl": "https://www.reuters.com/technology/openai-unveils-next-gen-reasoning-model-2026-04-21/",
"pubDate": "Tue, 21 Apr 2026 14:22:00 GMT",
"sourceName": "Reuters",
"description": "The new model targets regulated industries with improved factual accuracy and audit logs.",
"guid": "CBMiU2h0dHBzOi8vd3d3LnJldXRlcnMuY29t..."
}

📋 Output Data Schema

Every article row includes these fields:

FieldTypeDescription
positioninteger1‑based rank in the current run.
keywordstringThe search keyword used.
titlestringArticle headline (HTML stripped).
linkstringURL as returned by the RSS feed (often a Google News redirect).
articleUrlstringBest‑effort publisher URL unwrapped from Google's redirect.
pubDatestringPublication timestamp from the feed.
sourceNamestringPublishing outlet (e.g. Reuters, Bloomberg).
descriptionstringSnippet / summary text (HTML stripped).
guidstringStable feed ID — used for deduplication.

▶️ How to Use

  1. Run on Apify Console: Open the Google News Scraper Actor page, click Try for free, enter your keyword and numberOfResults, and press Start. Download results from the Dataset tab or RESULTS_CSV in Storage.
  2. Via API: Call the Actor with the Apify API using ApifyClient to fetch the dataset on run completion.
  3. Via CLI: Run with the Apify CLI:
    $apify call scrapeio/google-news-scraper --input='{"keyword":"climate policy","numberOfResults":500}'

🔗 API Example (JavaScript)

const { ApifyClient } = require('apify-client');
const client = new ApifyClient({ token: 'YOUR_APIFY_TOKEN' });
const run = await client.actor('scrapeio/google-news-scraper').call({
keyword: 'semiconductor supply chain',
numberOfResults: 100,
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(`Fetched ${items.length} unique headlines`);
items.slice(0, 5).forEach(a => console.log(`${a.pubDate}${a.sourceName}: ${a.title}`));

📈 Why Use This Google News Scraper?

  • Speed & Automation: Pull hundreds of fresh headlines per keyword in seconds — no manual searching, no copy‑paste, no scraper maintenance.
  • Scale & Reliability: Smart multi‑window RSS merging plus rel="next" pagination delivers depth, while GUID/link deduplication keeps rows unique.
  • Excel‑Ready Exports: RESULTS_CSV uses UTF‑8 BOM, quoted fields, and CRLF line endings — opens cleanly in Excel, Google Sheets, or Power BI.
  • Publisher URL Resolution: The Actor unwraps Google's redirect links when possible, giving you the direct publisher URL for click‑tracking, sentiment pipelines, and citation.
  • Scheduler & Webhooks: Combine with the Apify scheduler to run every hour, and webhook the new Dataset into Slack, HubSpot, or your data warehouse.
  • No API Quotas: Skip Google Cloud's paid News API, enterprise contracts, and throttles — this Actor uses the public RSS layer.

❓ FAQ

Q: Is it legal to scrape Google News? Google News RSS feeds are public. Storing headlines and links is widely considered acceptable for monitoring and research — but full‑text republishing is governed by publisher copyright and Google's terms. Always respect source outlets' robots and ToS.

Q: Does this call the paid Google News API? No. This Actor fetches the public news.google.com/rss/search feed. There is no Google Cloud project, billing account, or News API key involved.

Q: Why do I get fewer articles than numberOfResults? Google's index may not have that many unique stories for your keyword. Check OUTPUT.meta.stoppedReason: partial_no_more_feed means the feed was exhausted; completed means the target was met.

Q: Can I change country or language? The current release locks locale to hl=en-US / gl=US for reproducible RSS URLs. Custom locales are on the roadmap — contact the maintainer for priority.

Q: What output formats are supported? Dataset exports to CSV, JSON, Excel, XML, and HTML. An Excel‑friendly RESULTS_CSV and a full RESULTS_JSON are also written to the key‑value store.

Q: Does it handle rate limits? Yes — requests are spaced and retried on transient HTTP errors. You can also add Apify Proxy if needed for high‑frequency scheduled runs.

Q: Is articleUrl always the publisher URL? When Google wraps the publisher URL in a redirect, the Actor attempts to unwrap it. When unwrapping fails, articleUrl falls back to the RSS link (a Google News URL).

Q: Can I run this on a schedule? Yes. Use the Apify scheduler to trigger the Actor every hour, day, or custom cron, and webhook new items into Slack, email, or a database.

📣 Start Monitoring Google News Today

Run the Google News Scraper on Apify now → and turn any keyword into a clean, deduplicated stream of headlines, sources, and links.


Combine with these Apify Actors to build full media‑intelligence pipelines:


Built by ScrapeIO on Apify.

FAQ

How do I scrape Google News articles? Pass keyword and numberOfResults to scrapeio/google-news-scraper. Pin it in Cursor MCP.

Can I run this Actor with Apify AI? Yes. In Apify Console, type a long request in Store search or the dashboard Ask Apify AI widget. Apify AI matches this README, fills the input schema, asks you to confirm, then returns the dataset. This Actor uses limited permissions (not full-account access) and is not a rental Actor, so it is eligible for Apify AI.

Does this work from Cursor, Claude, or ChatGPT? Yes. Connect the Apify MCP server. Agents call search-actors, fetch-actor-details, and call-actor. Ranking uses the same signals as Store search and Actor quality score: a clear README plus documented input and output schemas.

Powered by AdScrape.