GDELT Global News Monitoring Scraper avatar

GDELT Global News Monitoring Scraper

Pricing

from $6.80 / 1,000 results

Go to Apify Store
GDELT Global News Monitoring Scraper

GDELT Global News Monitoring Scraper

Monitor worldwide news coverage by keyword, country and language. Scrape matching articles with title, link, source domain, source country, language, publish time and lead image. Great for brand and PR monitoring, media research and geo signals. Export to JSON, CSV or Excel.

Pricing

from $6.80 / 1,000 results

Rating

5.0

(1)

Developer

Scrapers Lat

Scrapers Lat

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

12 hours ago

Last modified

Share

GDELT Global News Monitoring Scraper

GDELT Global News Monitoring Scraper

Here is one real result, with every field the actor returns:

{
"title": "HubSpot Aktie : 7 . 000 Kunden verfehlen Ziel",
"url": "https://www.stock-world.de/hubspot-aktie-7-000-kunden-verfehlen-ziel/",
"urlMobile": null,
"domain": "stock-world.de",
"sourceCountry": "Germany",
"language": "German",
"seenDate": "2026-08-14T05:30:00.000Z",
"seenDateRaw": "20260814T053000Z",
"socialImage": "https://stock-world-de.b-cdn.net/wp-content/uploads/ai-imgs/US4435731009-2026-08-13-19-10-10.png",
"summary": "HubSpot steigert Umsatz und Gewinn im Q2, verfehlt jedoch das Kundenwachstumsziel und senkt die Jahresprognose. Der Aktienkurs fällt deutlich.",
"author": "Dieter Jaworski",
"publishedTime": "2026-08-13T17:10:09+00:00",
"keywords": null,
"sourceName": "Stock World",
"source": "GDELT",
"aiSummary": "HubSpot reported an increase in revenue and profit for Q2, but fell short of its customer growth target and has lowered its annual forecast, resulting in a significant drop in its stock price.",
"aiSentiment": "negative",
"aiSentimentScore": -0.6,
"aiTopics": ["HubSpot", "Aktie", "Kundenwachstum", "Umsatz", "Gewinn", "Jahresprognose"],
"aiCategory": "business",
"observedAt": "2026-08-14T05:55:49.219Z"
}

The most complete GDELT news monitoring scraper available. It returns every field the GDELT article feed exposes per story, from title, source domain, country and language to the lead image and seen date, plus per-article enrichment (summary, author, publish time, outlet name) and optional AI add-ons, and gives you filters for keyword, country, language, timespan and sort order to target exactly the coverage you need.

📥 Input · 📤 Output · 💰 Pricing · ▶️ Examples

Apify Coverage Output Billing

Table of contents

What it does

The actor queries the GDELT article feed for your keyword within the country, language, timespan and sort order you set, and writes one normalized record per article to the run's dataset. When enrichArticles is on (default), each article page is fetched to add a summary, author, publish time, keywords and outlet name from its metadata. Dates are normalized to ISO 8601, missing source values are returned as null, and optional AI add-ons (summary, sentiment/tone, topics and category) run per article when enabled on a paid plan.

Data covers worldwide online news monitored by GDELT across many countries and languages. The GDELT article feed returns up to 250 articles per query.

Quickstart

Open the actor, paste this into the input, and press Run. It returns 3 recent articles mentioning artificial intelligence, with enrichment and all AI add-ons on.

{
"query": "artificial intelligence",
"maxArticles": 3,
"timespan": "1d",
"enrichArticles": true,
"withSummary": true,
"withSentiment": true,
"withTopics": true,
"sort": "DateDesc"
}

Wrap a phrase in double quotes to match it exactly. A query is required. maxArticles defaults to 10 and is capped at 250 by the source feed.

Input reference

FieldTypeRequiredDefaultDescription
querystringyesartificial intelligenceKeyword or phrase to monitor across worldwide news. Wrap a phrase in double quotes to match it exactly.
maxArticlesintegerno10Maximum articles to collect (up to 250 per run).
countrystringno(any)Keep only articles from outlets in one country. GDELT country name or code, for example US, UK, France, Brazil.
languagestringno(any)Keep only articles in one language, for example english, spanish, portuguese, chinese.
timespanstringno3dHow far back to look, ending now. Number plus unit: min, h, d, w, m (for example 3d, 12h, 1w).
sortenumnoDateDescDateDesc (newest), DateAsc (oldest), ToneDesc (most positive), ToneAsc (most negative), HybridRel (relevance).
enrichArticlesbooleannotrueAdd summary, author, publish time, keywords and outlet name from each article's page. Turn off for a faster headline-only run.
withSummarybooleannofalsePaid add-on. Adds aiSummary. Charged only on usable output. Requires a paid Apify plan.
withSentimentbooleannofalsePaid add-on. Adds aiSentiment and aiSentimentScore. Charged only on usable output. Requires a paid Apify plan.
withTopicsbooleannofalsePaid add-on. Adds aiTopics and aiCategory. Charged only on usable output. Requires a paid Apify plan.

Filters combine with logical AND. Empty filters are ignored.

Output reference

One dataset item per article. Types: string, number, string[], or null when the source value is absent.

FieldTypeDescription
titlestringArticle headline.
urlstringArticle URL.
urlMobilestringMobile URL when provided, else null.
domainstringSource domain.
sourceCountrystringCountry of the outlet.
languagestringLanguage of the article.
seenDatestringISO 8601 time GDELT first saw the article.
seenDateRawstringRaw GDELT seen-date string.
socialImagestringLead/social image URL, or null.
summarystringSummary from the article metadata (enrichment), or null.
authorstringAuthor from the article metadata, or null.
publishedTimestringPublish time from the article metadata, or null.
keywordsstring[]Keywords from the article metadata, or null.
sourceNamestringOutlet name from the article metadata, or null.
sourcestringAlways GDELT.
aiSummarystringAI neutral summary. null unless withSummary produced output.
aiSentimentstringAI tone (positive, negative, neutral). null unless withSentiment produced output.
aiSentimentScorenumberAI tone score (roughly -1 to 1). null unless withSentiment produced output.
aiTopicsstring[]AI topic/keyword tags. null unless withTopics produced output.
aiCategorystringAI broad news category. null unless withTopics produced output.
observedAtstringISO 8601 timestamp of when the record was collected.
errorstringOn a failed run, a single item with a populated error field is written instead.

Example output record

Real record from a live run (input {"query": "artificial intelligence", "maxArticles": 3, "timespan": "1d", "enrichArticles": true, "withSummary": true, "withSentiment": true, "withTopics": true}):

{
"title": "HubSpot Aktie : 7 . 000 Kunden verfehlen Ziel",
"url": "https://www.stock-world.de/hubspot-aktie-7-000-kunden-verfehlen-ziel/",
"domain": "stock-world.de",
"sourceCountry": "Germany",
"language": "German",
"seenDate": "2026-08-14T05:30:00.000Z",
"seenDateRaw": "20260814T053000Z",
"socialImage": "https://stock-world-de.b-cdn.net/wp-content/uploads/ai-imgs/US4435731009-2026-08-13-19-10-10.png",
"summary": "HubSpot steigert Umsatz und Gewinn im Q2, verfehlt jedoch das Kundenwachstumsziel und senkt die Jahresprognose. Der Aktienkurs fällt deutlich.",
"author": "Dieter Jaworski",
"publishedTime": "2026-08-13T17:10:09+00:00",
"sourceName": "Stock World",
"source": "GDELT",
"aiSummary": "HubSpot reported an increase in revenue and profit for Q2, but fell short of its customer growth target and has lowered its annual forecast, resulting in a significant drop in its stock price.",
"aiSentiment": "negative",
"aiSentimentScore": -0.6,
"aiTopics": ["HubSpot", "Aktie", "Kundenwachstum", "Umsatz", "Gewinn", "Jahresprognose"],
"aiCategory": "business",
"observedAt": "2026-08-14T05:55:49.219Z"
}

Run via API and CLI

Start a run and wait for it to finish, then read the dataset. Replace <TOKEN> with your Apify API token.

Run synchronously and get dataset items in one call:

curl -X POST "https://api.apify.com/v2/acts/scrapers_lat~gdelt-news-events-scraper/run-sync-get-dataset-items?token=<TOKEN>" \
-H "Content-Type: application/json" \
-d '{"query":"\"Tesla recall\"","timespan":"7d","maxArticles":25}'

Start a run asynchronously:

curl -X POST "https://api.apify.com/v2/acts/scrapers_lat~gdelt-news-events-scraper/runs?token=<TOKEN>" \
-H "Content-Type: application/json" \
-d '{"query":"artificial intelligence","country":"US","language":"english","maxArticles":250}'

Apify CLI:

apify call scrapers_lat/gdelt-news-events-scraper \
--input '{"query":"climate policy","sort":"ToneAsc","maxArticles":50}'

Fetch results

Every run writes to a dataset. Fetch items as JSON, CSV, or Excel by changing format:

# JSON
curl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&clean=true&format=json"
# CSV
curl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&clean=true&format=csv"
# Paginate large datasets
curl "https://api.apify.com/v2/datasets/<DATASET_ID>/items?token=<TOKEN>&offset=1000&limit=1000"

<DATASET_ID> is returned as defaultDatasetId in the run object. Use offset and limit to page through large result sets. clean=true drops empty and internal fields.

Billing and limits

  • Pay per result. You are charged per article record returned (result event). See the pricing tab for the current per-result price.
  • Optional AI add-ons (ai_summary, ai_sentiment, ai_topics) are billed separately and only when they produce usable output, and only on paid Apify plans.
  • No charge on failure. If a run errors, the actor writes a single item with a populated error field and does not charge for it. Empty runs cost nothing.
  • Spend cap respected. Set maxTotalChargeUsd on the run; once reached, the actor stops emitting and charging further billable results.
  • Free Apify plans are capped at 10 records per run. Upgrade for higher maxArticles (up to 250, the source feed ceiling).

FAQ and troubleshooting

A run returned 0 records. Why? No articles matched the keyword within the timespan and country/language filters. Widen timespan, remove country/language, or broaden the keyword. Zero-result runs are not charged.

How do I match an exact phrase? Wrap it in double quotes, for example "Tesla recall". Without quotes the words are matched individually.

Why are author or keywords null? They come from the article page metadata during enrichment. If enrichArticles is off, or the page did not expose them, they stay null. The actor never invents values.

How far back can I search? Set timespan with a number and unit (min, h, d, w, m). The GDELT feed returns at most 250 articles per query.

Do the AI fields cost extra? Yes. They are opt-in paid add-ons, billed only when they return usable output, and disabled on free Apify plans.

Is this an official GDELT tool? No. This actor is independent and has no affiliation with the GDELT Project. It reads only data that is publicly available through GDELT. Use it in accordance with the GDELT terms of use.

More scrapers at scrapers.lat

Built and maintained by scrapers.lat, where we publish scrapers for US and Latin American public platforms: company registries, government data, finance, e-commerce and more. Browse the catalog or request a custom scraper at scrapers.lat.


Independent tool, not affiliated with the GDELT Project. Accesses only publicly available data. Use in accordance with the GDELT terms of use.