MarketWatch Scraper avatar

MarketWatch Scraper

Pricing

from $1.00 / 1,000 records

Go to Apify Store
MarketWatch Scraper

MarketWatch Scraper

Extract MarketWatch top stories and real-time bulletins as structured news: headline, summary, byline, an ISO 8601 publication time, and the company tickers each story mentions. Export data, run via API, schedule and monitor runs, or integrate with other tools.

Pricing

from $1.00 / 1,000 records

Rating

5.0

(1)

Developer

Public Money

Public Money

Maintained by Apify

Actor stats

1

Bookmarked

4

Total users

3

Monthly active users

21 hours ago

Last modified

Categories

Share

MarketWatch is the equities and macro desk in this fleet, not a crypto feed, and its bulletins carry market-moving headlines within minutes. This Actor turns them into structured records: headline, summary, byline and a real ISO 8601 timestamp, each tagged with the company tickers the story actually names, resolved against the SEC's own ticker file.

What it does

  • Covers equities and macro rather than crypto, so it pairs with the Yahoo, Nasdaq and CNBC Actors rather than the coin feeds.
  • Reads the bulletins feed, which carries short market-moving headlines within minutes rather than finished articles.
  • Resolves company tickers against the SEC company ticker file, so mentionedTickers maps to real registrants rather than guesses.
  • Tags which assets each story is about in mentionedTickers, resolved against a live CoinGecko coin list and the SEC company ticker file rather than a hardcoded table, with a guard so an ordinary word like "optimism" does not ship OP as a mention.
  • Converts the feed's RFC 822 date to a real ISO 8601 datePublished, so stories sort and filter correctly instead of arriving as "Sat, 06 Sep 2026 05:30:00 GMT" strings.
  • Returns the summary body in articleBody, decoded from the feed's HTML entities, so a headline arrives with enough context to classify without a second fetch.
  • Carries the feed's own guid, so a scheduled run can drop stories it has already seen instead of reprocessing the whole feed.

Use cases

You need toHow this Actor does it
Alert on a holdingFilter mentionedTickers against your portfolio and push to Slack on a match
Watch the tapeRead the bulletins feed on a short schedule for market-moving headlines
Join a headline to a price moveRun the Yahoo Finance or Nasdaq Actor on the same schedule and match on ticker
Build a news timeline for a companyAccumulate runs and group by mentionedTickers and datePublished
Feed a finance agentCall the Actor over MCP and let the model read the day's headlines
Deduplicate a feedKey your store on guid and only process what is new

Quick start

  1. Click Try for free.
  2. Leave Feeds on topstories and bulletins, the two that are still updated. marketpulse and realtimeheadlines are served but no longer maintained.
  3. Set Articles per feed to control how many stories come back.
  4. Click Start. Rows appear within seconds.
  5. Export as JSON, CSV, Excel or XML, or read the dataset over the API.

Input

FieldTypeDefaultWhat it controls
feedsarraytopstories, bulletinsWhich MarketWatch feeds to read. topstories and bulletins are live; marketpulse and realtimeheadlines are still served but Dow Jones stopped adding to them in 2025
articlesPerFeedinteger50How many stories each feed returns, counted from the newest
maxItemsinteger0Caps how many records are written. 0 writes them all
{
"articlesPerFeed": 50,
"maxItems": 0
}

Output

One dataset item per story. Fields the feed does not carry for a story are dropped rather than returned as null, so a story with no byline carries no author.

Field groupFields
Storystatus, headline, articleBody, url, guid
Attributionpublisher, author, categories, feed
AssetsmentionedTickers
TimingdatePublished, scrapedAt
{
"status": "ok",
"headline": "Fed holds rates, signals one cut before year end",
"articleBody": "The decision was unanimous, with the statement pointing to cooling wage growth.",
"publisher": "MarketWatch",
"author": "Greg Robb",
"categories": [],
"feed": "bulletins",
"mentionedTickers": [
"SPY"
],
"datePublished": "2026-09-06T05:30:00.000Z",
"scrapedAt": "2026-09-06T06:02:11.418Z",
"url": "https://www.marketwatch.com/story/fed-holds-rates-signals-one-cut"
}

Integrations

Run it over the API and get the rows back in one call:

curl -X POST "https://api.apify.com/v2/acts/publicmoney~marketwatch-scraper/run-sync-get-dataset-items?token=YOUR_TOKEN" \
-H "Content-Type: application/json" \
-d '{"articlesPerFeed": 50, "maxItems": 0}'

From Python:

from apify_client import ApifyClient
client = ApifyClient("YOUR_TOKEN")
run = client.actor("publicmoney/marketwatch-scraper").call(run_input={"articlesPerFeed": 50, "maxItems": 0})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item["headline"])

Give an AI agent the Actor over MCP:

{
"mcpServers": {
"apify": {
"url": "https://mcp.apify.com/?actors=publicmoney/marketwatch-scraper"
}
}
}

Schedules run it on any cron, webhooks fire when a run finishes, and platform integrations push the dataset to Google Sheets, Slack, Airtable, Zapier or your own endpoint.

Cost

Pay per event, so you pay for records rather than compute time.

EventFree tierTop volume tier
News article$0.002$0.0007
Actor start$0.00005 per GBSame

A record that returned no data is published as a failure row and is never charged. Six volume tiers apply, so the per-record price falls with monthly volume.

Troubleshooting

IssueSolution
Every story returns failedThe MarketWatch feed was unreachable on that run. It is a public endpoint, so this is usually transient. Check the log and rerun.
marketpulse or realtimeheadlines returns nothing newDow Jones stopped adding to those two in 2025. They still serve, so old items come back, but use topstories and bulletins for current coverage.
Stories carry no categoriesMarketWatch does not tag its feed items the way the crypto publications do. Use feed and mentionedTickers to route instead.
mentionedTickers is empty on a story that clearly names an assetThe ticker index is resolved live and a story that names a company only in prose, with no cashtag and no bracketed symbol, will not match. An index that fails to load leaves the field empty rather than failing the run, so check the log.
A scheduled run returns stories I already haveThe feed carries a rolling window, so each run sees recent stories again. Deduplicate on guid, which is stable per story.
articleBody is a summary, not the full articleThat is what the feed publishes. This Actor does not fetch the article page, so you get the headline, the summary and the link.

FAQ

Does MarketWatch have an API?

No public API. MarketWatch is a Dow Jones title and its data is sold through Dow Jones, while the syndication feeds are open, which is what this Actor reads.

Is this a stock news API?

It is the closest thing to one here: MarketWatch headlines with the company tickers each story names, resolved against the SEC ticker file. For per-symbol news pages use the Yahoo Finance, Nasdaq or Finviz Actor in this fleet.

Which feeds are still updated?

topstories and bulletins. marketpulse and realtimeheadlines are still served but Dow Jones stopped adding to them in 2025, so treat them as an archive.

How does it know which assets a story is about?

It resolves cashtags and bracketed symbols in the headline and body against a live CoinGecko coin list and the SEC company ticker file. Both are fetched per run rather than hardcoded, so new listings resolve without an Actor update, and there is a guard against common words that happen to be tickers.

Does it return the full article text?

No. It returns what the feed publishes: headline, summary body, author, categories, link and date. Fetching and republishing full articles is a copyright question, so the Actor gives you the metadata and the URL.

Do I need a MarketWatch API key?

No. MarketWatch publishes its feed openly and this Actor reads it, so there is no MarketWatch credential anywhere. You need an Apify token to call the Actor over the API.

Can I get this data in Python?

Yes, with the apify-client package as shown above. It returns parsed JSON, so there is no HTML or response handling on your side.

Can I get the data into Excel or Google Sheets?

Yes. Export the dataset as XLSX or CSV, or connect the Google Sheets integration so each run appends to a sheet.

Can an AI agent call this Actor?

Yes. Add it to an MCP client with the config above and the model can request what it needs on its own. Every record is flat JSON with named fields, so no post-processing is needed.

This Actor reads MarketWatch's own public feed, which the publisher publishes for syndication, and returns headline, summary, byline and link rather than full article text. Republishing the text of an article is a copyright question and is your responsibility, so take your own legal advice for how you use it.

Changelog

  • 0.0.2 Added the bulletins feed and mentioned-ticker resolution against the SEC ticker file.
  • 0.0.1 First release. MarketWatch headlines as structured records.

Feedback

Found a field MarketWatch publishes that this Actor misses, or an input it rejects? Open an issue on the Issues tab with the input and what you expected. A daily test runs every Actor in the fleet against live sources, so parser fixes ship fast.