Google News Scraper (RSS search by keyword, topic, site) avatar

Google News Scraper (RSS search by keyword, topic, site)

Pricing

from $0.30 / 1,000 articles

Go to Apify Store
Google News Scraper (RSS search by keyword, topic, site)

Google News Scraper (RSS search by keyword, topic, site)

Google News Scraper returns headlines from Google News RSS search and topic feeds for any keyword, site: or when: query, and resolves each article's real publisher URL — one row per article.

Pricing

from $0.30 / 1,000 articles

Rating

0.0

(0)

Developer

Murat Uzun

Murat Uzun

Maintained by Community

Actor stats

0

Bookmarked

1

Total users

1

Monthly active users

2 days ago

Last modified

Categories

Share

What is Google News Scraper?

Google News Scraper is an Apify Actor that reads Google News RSS search and section-topic feeds for any list of keywords, site: searches or when:Nd recency queries, and returns one clean row per article — title, source, publish date and the Google News link. No API key, no browser, no anti-bot wall: Google's own public news.google.com/rss/* feeds do the work. When Resolve publisher URLs is on, the Actor also follows each article's news.google.com/rss/articles/... redirect link and resolves the publisher's real URL (see "Publisher URL resolution" below) instead of leaving you with a Google-branded link nobody wants to click.

What data does Google News Scraper extract?

Google News Scraper extracts 11 fields per article:

FieldTypeDescription
querystringThe search query, or topic:NAME, this article came from
titlestringHeadline as shown by Google News (often "Headline - Publisher")
googleUrlstringThe news.google.com/rss/articles/... link from the feed
publisherUrlstring|nullThe article's real URL on the publisher's own site, resolved best-effort
source, sourceUrlstringPublisher name and homepage, from the feed's <source> tag
publishedAtstringPublish date/time (ISO 8601 UTC)
snippetstring|nullPlain text derived from the feed's description
guidstringThe feed's stable item id, used to dedupe articles across overlapping queries
error, scrapedAtstringPer-query failure note and check timestamp

How to use Google News Scraper

  1. Enter one or more Search queries — plain keywords, or Google operators like site:reuters.com AI or "exact phrase". Add Section topics (TECHNOLOGY, BUSINESS, WORLD…) if you also want whole-section feeds.
  2. Set Only articles from the last N days if you want a recency window (adds Google's when:Nd operator to every query).
  3. Leave Resolve publisher URLs on to get the real article link, click Start, then export as JSON, CSV, Excel or HTML.

Example input

{
"queries": ["web scraping", "site:reuters.com AI"],
"topics": ["TECHNOLOGY"],
"sinceDays": 7,
"resolvePublisherUrls": true,
"maxItemsPerQuery": 50
}

Example output

{
"query": "web scraping",
"title": "EDPB web scraping guidelines for AI: Making the impossible possible? - Reed Smith LLP",
"googleUrl": "https://news.google.com/rss/articles/CBMi3AFBVV95cUxOTUxq...?oc=5",
"publisherUrl": "https://www.reedsmith.com/our-insights/blogs/technology-law-dispatch/102nbqu/edpb-web-scraping-guidelines-for-ai-making-the-impossible-possible/",
"source": "Reed Smith LLP",
"sourceUrl": "https://www.reedsmith.com",
"publishedAt": "2026-07-14T07:00:00.000Z",
"snippet": "EDPB web scraping guidelines for AI: Making the impossible possible? Reed Smith LLP",
"guid": "CBMi3AFBVV95cUxOTUxq...",
"error": null,
"scrapedAt": "2026-09-13T00:00:00.000Z"
}

Input parameters

ParameterTypeDefaultDescription
queriesarray["web scraping"]Search terms/operators, one Google News search feed per entry
topicsarray[]Section topics: WORLD, NATION, BUSINESS, TECHNOLOGY, ENTERTAINMENT, SCIENCE, SPORTS, HEALTH
language / countrystringen / USInterface language and edition (builds hl/gl/ceid)
sinceDaysinteger0Adds when:Nd to searches; 0 disables it (topics ignore it)
maxItemsPerQueryinteger100Cap per query/topic (Google itself rarely returns more than ~100)
resolvePublisherUrlsbooleantrueResolve the real publisher URL (2 extra requests/article)
maxConcurrencyinteger5Parallel feed fetches, and parallel URL resolutions (1-20)

Pricing

Google News Scraper uses pay-per-event pricing: $0.0005 per article result, i.e. $0.50 per 1,000 articles, plus a negligible actor-start fee. A single feed request typically returns dozens of articles for a fraction of a cent; resolving publisher URLs adds two extra requests per article but no extra dataset charge. Set Maximum cost per run and the Actor trims the result list to what the budget covers instead of overspending.

Google News Scraper vs. manual Google News browsing

Reading Google News by hand means opening dozens of tabs and copy-pasting headlines one at a time, and every article link is a Google-branded redirect you cannot drop into a report. This Actor takes a whole list of keywords, sites and topics in one run, returns a flat structured dataset, resolves the real publisher link, and can be scheduled to catch new coverage the moment it appears.

Using Google News Scraper with AI agents and MCP

Google News Scraper is pay-per-event with limited permissions — the two requirements for an Actor to be callable through the Apify MCP server at mcp.apify.com. An agent passes queries or topics and gets back one structured row per article — headline, source, publish date, and (best-effort) the real publisher link — ready to feed into a summarizer, a lead-monitoring workflow or a competitor-tracking pipeline. The same run works from n8n, Make, Zapier and LangChain through Apify's integrations.

FAQ

How does publisher URL resolution work? Google's feed link is never the publisher's own URL — it is a news.google.com/rss/articles/<id> redirect that answers a 302 pointing back to itself, then a client-rendered page. This Actor fetches that page, extracts the data-n-a-id/data-n-a-ts/data-n-a-sg values Google embeds in it, and posts them to Google's internal redirect-resolution endpoint to get the real URL back. This is an undocumented internal API, verified live but not published by Google — it can break without notice if Google changes the page or endpoint shape, in which case affected rows simply get publisherUrl: null and keep everything else.

Why is publisherUrl null on some rows? Either Resolve publisher URLs was off, the resolution request failed or timed out, or Google's page/endpoint shape changed since this Actor was last verified. The row still ships with its Google News link, title, source and date.

Does when:Nd apply to section topics? No — topic feeds (topics[]) are always "latest for that section"; the recency filter only applies to queries[] searches, where it is a genuine Google search operator (verified live: site:reuters.com AI when:7d returned only reuters.com articles from the prior week).

What are the limitations? Google's feeds return at most roughly 100 items per request with no further pagination, so a very broad query is naturally capped. Google News' own copyright notice restricts the feed to personal, non-commercial feed-reader use; treat bulk/commercial reuse of the underlying articles as your responsibility. A failing query still produces one row, with the reason in error.

Is this legal to run? The RSS feeds are Google's own public syndication format. Downstream use of the article text/links is subject to Google's feed terms and each publisher's own copyright — this Actor only extracts metadata and links, it does not copy article bodies.

Can I export to CSV or Excel? Yes, from the Output tab or the API, with a ready-made Overview view.

Part of the webdatatools web-intelligence suite — every Actor is pay-per-event, reads public data without a login, and returns one clean row per entity:

Browse the whole suite at webdatatools, or call ten of these Actors straight from Claude, Cursor or Cline with the webdatatools MCP server.

Website & domain intelligence

Content for AI, LLMs and RAG

Search, video and social

Leads, jobs and company data

Developer, app and research data

Support and feedback

Found a topic Google added, a feed shape that changed, or a resolution bug? Open an issue on the Issues tab.