GDELT Global News Monitoring Scraper
Pricing
from $6.80 / 1,000 results
GDELT Global News Monitoring Scraper
Monitor worldwide news coverage by keyword, country and language. Scrape matching articles with title, link, source domain, source country, language, publish time and lead image. Great for brand and PR monitoring, media research and geo signals. Export to JSON, CSV or Excel.
Pricing
from $6.80 / 1,000 results
Rating
5.0
(1)
Developer
Scrapers Lat
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
12 hours ago
Last modified
Categories
Share
GDELT Global News Monitoring Scraper
Here is one real result, with every field the actor returns:
{"title": "HubSpot Aktie : 7 . 000 Kunden verfehlen Ziel","url": "https://www.stock-world.de/hubspot-aktie-7-000-kunden-verfehlen-ziel/","urlMobile": null,"domain": "stock-world.de","sourceCountry": "Germany","language": "German","seenDate": "2026-08-14T05:30:00.000Z","seenDateRaw": "20260814T053000Z","socialImage": "https://stock-world-de.b-cdn.net/wp-content/uploads/ai-imgs/US4435731009-2026-08-13-19-10-10.png","summary": "HubSpot steigert Umsatz und Gewinn im Q2, verfehlt jedoch das Kundenwachstumsziel und senkt die Jahresprognose. Der Aktienkurs fällt deutlich.","author": "Dieter Jaworski","publishedTime": "2026-08-13T17:10:09+00:00","keywords": null,"sourceName": "Stock World","source": "GDELT","aiSummary": "HubSpot reported an increase in revenue and profit for Q2, but fell short of its customer growth target and has lowered its annual forecast, resulting in a significant drop in its stock price.","aiSentiment": "negative","aiSentimentScore": -0.6,"aiTopics": ["HubSpot", "Aktie", "Kundenwachstum", "Umsatz", "Gewinn", "Jahresprognose"],"aiCategory": "business","observedAt": "2026-08-14T05:55:49.219Z"}
The most complete GDELT news monitoring scraper available. It returns every field the GDELT article feed exposes per story, from title, source domain, country and language to the lead image and seen date, plus per-article enrichment (summary, author, publish time, outlet name) and optional AI add-ons, and gives you filters for keyword, country, language, timespan and sort order to target exactly the coverage you need.
📥 Input · 📤 Output · 💰 Pricing · ▶️ Examples
Table of contents
- What it does
- Quickstart
- Input reference
- Output reference
- Example output record
- Run via API and CLI
- Fetch results
- Billing and limits
- FAQ and troubleshooting
What it does
The actor queries the GDELT article feed for your keyword within the country, language, timespan and sort order you set, and writes one normalized record per article to the run's dataset. When enrichArticles is on (default), each article page is fetched to add a summary, author, publish time, keywords and outlet name from its metadata. Dates are normalized to ISO 8601, missing source values are returned as null, and optional AI add-ons (summary, sentiment/tone, topics and category) run per article when enabled on a paid plan.
Data covers worldwide online news monitored by GDELT across many countries and languages. The GDELT article feed returns up to 250 articles per query.
Quickstart
Open the actor, paste this into the input, and press Run. It returns 3 recent articles mentioning artificial intelligence, with enrichment and all AI add-ons on.
{"query": "artificial intelligence","maxArticles": 3,"timespan": "1d","enrichArticles": true,"withSummary": true,"withSentiment": true,"withTopics": true,"sort": "DateDesc"}
Wrap a phrase in double quotes to match it exactly. A query is required. maxArticles defaults to 10 and is capped at 250 by the source feed.
Input reference
| Field | Type | Required | Default | Description |
|---|---|---|---|---|
query | string | yes | artificial intelligence | Keyword or phrase to monitor across worldwide news. Wrap a phrase in double quotes to match it exactly. |
maxArticles | integer | no | 10 | Maximum articles to collect (up to 250 per run). |
country | string | no | (any) | Keep only articles from outlets in one country. GDELT country name or code, for example US, UK, France, Brazil. |
language | string | no | (any) | Keep only articles in one language, for example english, spanish, portuguese, chinese. |
timespan | string | no | 3d | How far back to look, ending now. Number plus unit: min, h, d, w, m (for example 3d, 12h, 1w). |
sort | enum | no | DateDesc | DateDesc (newest), DateAsc (oldest), ToneDesc (most positive), ToneAsc (most negative), HybridRel (relevance). |
enrichArticles | boolean | no | true | Add summary, author, publish time, keywords and outlet name from each article's page. Turn off for a faster headline-only run. |
withSummary | boolean | no | false | Paid add-on. Adds aiSummary. Charged only on usable output. Requires a paid Apify plan. |
withSentiment | boolean | no | false | Paid add-on. Adds aiSentiment and aiSentimentScore. Charged only on usable output. Requires a paid Apify plan. |
withTopics | boolean | no | false | Paid add-on. Adds aiTopics and aiCategory. Charged only on usable output. Requires a paid Apify plan. |
Filters combine with logical AND. Empty filters are ignored.
Output reference
One dataset item per article. Types: string, number, string[], or null when the source value is absent.
| Field | Type | Description |
|---|---|---|
title | string | Article headline. |
url | string | Article URL. |
urlMobile | string | Mobile URL when provided, else null. |
domain | string | Source domain. |
sourceCountry | string | Country of the outlet. |
language | string | Language of the article. |
seenDate | string | ISO 8601 time GDELT first saw the article. |
seenDateRaw | string | Raw GDELT seen-date string. |
socialImage | string | Lead/social image URL, or null. |
summary | string | Summary from the article metadata (enrichment), or null. |
author | string | Author from the article metadata, or null. |
publishedTime | string | Publish time from the article metadata, or null. |
keywords | string[] | Keywords from the article metadata, or null. |
sourceName | string | Outlet name from the article metadata, or null. |
source | string | Always GDELT. |
aiSummary | string | AI neutral summary. null unless withSummary produced output. |
aiSentiment | string | AI tone (positive, negative, neutral). null unless withSentiment produced output. |
aiSentimentScore | number | AI tone score (roughly -1 to 1). null unless withSentiment produced output. |
aiTopics | string[] | AI topic/keyword tags. null unless withTopics produced output. |
aiCategory | string | AI broad news category. null unless withTopics produced output. |
observedAt | string | ISO 8601 timestamp of when the record was collected. |
error | string | On a failed run, a single item with a populated error field is written instead. |
Example output record
Real record from a live run (input {"query": "artificial intelligence", "maxArticles": 3, "timespan": "1d", "enrichArticles": true, "withSummary": true, "withSentiment": true, "withTopics": true}):
{"title": "HubSpot Aktie : 7 . 000 Kunden verfehlen Ziel","url": "https://www.stock-world.de/hubspot-aktie-7-000-kunden-verfehlen-ziel/","domain": "stock-world.de","sourceCountry": "Germany","language": "German","seenDate": "2026-08-14T05:30:00.000Z","seenDateRaw": "20260814T053000Z","socialImage": "https://stock-world-de.b-cdn.net/wp-content/uploads/ai-imgs/US4435731009-2026-08-13-19-10-10.png","summary": "HubSpot steigert Umsatz und Gewinn im Q2, verfehlt jedoch das Kundenwachstumsziel und senkt die Jahresprognose. Der Aktienkurs fällt deutlich.","author": "Dieter Jaworski","publishedTime": "2026-08-13T17:10:09+00:00","sourceName": "Stock World","source": "GDELT","aiSummary": "HubSpot reported an increase in revenue and profit for Q2, but fell short of its customer growth target and has lowered its annual forecast, resulting in a significant drop in its stock price.","aiSentiment": "negative","aiSentimentScore": -0.6,"aiTopics": ["HubSpot", "Aktie", "Kundenwachstum", "Umsatz", "Gewinn", "Jahresprognose"],"aiCategory": "business","observedAt": "2026-08-14T05:55:49.219Z"}
Run via API and CLI
Start a run and wait for it to finish, then read the dataset. Replace <TOKEN> with your Apify API token.
Run synchronously and get dataset items in one call:
curl -X POST "https://api.apify.com/v2/acts/scrapers_lat~gdelt-news-events-scraper/run-sync-get-dataset-items?token=<TOKEN>" \-H "Content-Type: application/json" \-d '{"query":"\"Tesla recall\"","timespan":"7d","maxArticles":25}'
Start a run asynchronously:
curl -X POST "https://api.apify.com/v2/acts/scrapers_lat~gdelt-news-events-scraper/runs?token=<TOKEN>" \-H "Content-Type: application/json" \-d '{"query":"artificial intelligence","country":"US","language":"english","maxArticles":250}'
Apify CLI:
apify call scrapers_lat/gdelt-news-events-scraper \--input '{"query":"climate policy","sort":"ToneAsc","maxArticles":50}'
Fetch results
Every run writes to a dataset. Fetch items as JSON, CSV, or Excel by changing format:
# JSONcurl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&clean=true&format=json"# CSVcurl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&clean=true&format=csv"# Paginate large datasetscurl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&offset=1000&limit=1000"
<DATASET_ID> is returned as defaultDatasetId in the run object. Use offset and limit to page through large result sets. clean=true drops empty and internal fields.
Billing and limits
- Pay per result. You are charged per article record returned (
resultevent). See the pricing tab for the current per-result price. - Optional AI add-ons (
ai_summary,ai_sentiment,ai_topics) are billed separately and only when they produce usable output, and only on paid Apify plans. - No charge on failure. If a run errors, the actor writes a single item with a populated
errorfield and does not charge for it. Empty runs cost nothing. - Spend cap respected. Set
maxTotalChargeUsdon the run; once reached, the actor stops emitting and charging further billable results. - Free Apify plans are capped at 10 records per run. Upgrade for higher
maxArticles(up to 250, the source feed ceiling).
FAQ and troubleshooting
A run returned 0 records. Why?
No articles matched the keyword within the timespan and country/language filters. Widen timespan, remove country/language, or broaden the keyword. Zero-result runs are not charged.
How do I match an exact phrase?
Wrap it in double quotes, for example "Tesla recall". Without quotes the words are matched individually.
Why are author or keywords null?
They come from the article page metadata during enrichment. If enrichArticles is off, or the page did not expose them, they stay null. The actor never invents values.
How far back can I search?
Set timespan with a number and unit (min, h, d, w, m). The GDELT feed returns at most 250 articles per query.
Do the AI fields cost extra? Yes. They are opt-in paid add-ons, billed only when they return usable output, and disabled on free Apify plans.
Is this an official GDELT tool? No. This actor is independent and has no affiliation with the GDELT Project. It reads only data that is publicly available through GDELT. Use it in accordance with the GDELT terms of use.
Related scrapers
- Google News Scraper: News headlines by keyword and country.
- Federal Register Document Scraper: US Federal Register documents.
- CFPB Complaints Scraper: US consumer finance complaints.
- Hacker News Scraper: Hacker News stories and comments.
- Reddit Search Scraper: Reddit posts by keyword.
- Company Hiring Signal Scraper: Hiring signals from company career pages.
More scrapers at scrapers.lat
Built and maintained by scrapers.lat, where we publish scrapers for US and Latin American public platforms: company registries, government data, finance, e-commerce and more. Browse the catalog or request a custom scraper at scrapers.lat.
Independent tool, not affiliated with the GDELT Project. Accesses only publicly available data. Use in accordance with the GDELT terms of use.
