Newscrusher.ai AI News Scraper avatar

Newscrusher.ai AI News Scraper

Pricing

from $3.00 / 1,000 scraped articles

Go to Apify Store
Newscrusher.ai AI News Scraper

Newscrusher.ai AI News Scraper

The fastest, most comprehensive, and most reliable AI News Scraper on the store. Scrape Newscrusher.ai to extract trending AI news from across most of the newspapers on the internet and extract AI-generated summaries and more from each article!

Pricing

from $3.00 / 1,000 scraped articles

Rating

0.0

(0)

Developer

Coding Doctor Omar

Coding Doctor Omar

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

5 hours ago

Last modified

Categories

Share

🤖 Newscrusher.ai AI News Scraper

actor logo

The fastest, most comprehensive, and most reliable AI News Scraper on the store. Scrape Newscrusher.ai to extract trending AI news from across most of the newspapers on the internet and extract AI-generated summaries and more from each article!

Leave a ReviewInput SchemaOutput SchemaPricingAPI DocsIssues/Open an Issue

What Data Can You Get?

Newscrusher.ai AI News Scraper scrapes a powerful AI news aggregator website called newscrusher.ai. The website contains a live news feed of all AI News articles from almost all newspaper websites you can imagine, updated 24/7. Example newspapers whose articles are included in Newscrusher.ai include Wall Street Journal, Tech Crunch, Quartz, Futurism, The Verge, OpenAI, The New Stack, Towards Data Science, Geeky Gadgets, AWS ML Blog, and more. For each article, you get the following data:

  1. id — A unique Newscrusher.ai identifier for the article.
  2. headline — The headline of the article.
  3. subHeadline — The sub headline of the article.
  4. publishedAt — The date and time the article was published on the original newspaper website, in UTC timezone and ISO format.
  5. publisherDomain — The domain of the website of the publisher newspaper.
  6. newspaper — The name of the publisher newspaper.
  7. articleUrl — The URL for the article on the original newspaper website.
  8. priorityScore — An importance score provided by Newscrusher.ai to indicate how important each article is.
  9. articleTags — A list of tags for the article (will be null if article summaries are disabled in the Actor input).
  10. articleSummary — An AI-generated summary for the article's content (will be null if article summaries are disabled in the Actor input).
  11. readTime — The estimated time it would take the average reader to read the full article (will be null if article summaries are disabled in the Actor input).
  12. scrapedAt — The date and time the article was scraped from Newscrusher.ai, in UTC timezone and ISO format.
  13. keyTakeaways — An AI-generated list of 5 key takeaways to know from the article (will be null if article summaries are disabled in the Actor input).

Example Actor Output

{
"id": "5c7a0cfb",
"headline": "AIR raises $50M to help companies vet the skills and add-ons AI agents use",
"subHeadline": "AIR's platform can discover agents running at a company, continuously vets any skills and add-ons they use, and blocks any unwanted behaviour.",
"publishedAt": "2026-09-01T15:45:51+00:00",
"publisherDomain": "techcrunch.com",
"newspaper": "Techcrunch",
"articleUrl": "https://techcrunch.com/2026/09/01/air-raises-50m-to-help-companies-vet-the-skills-and-add-ons-ai-agents-use/",
"priorityScore": 70,
"articleTags": [
"Agents",
"Enterprise Adoption"
],
"articleSummary": "AIR secured $50 million in funding to expand its platform for discovering, monitoring, and controlling AI agents within enterprise environments. The solution identifies agents running across company systems, continuously evaluates their capabilities and integrations, and prevents unauthorized or risky behaviors from executing.<br><br>The platform addresses growing concerns about AI agent proliferation in enterprises, where multiple agents may operate independently without centralized oversight. By providing visibility and control mechanisms, AIR enables organizations to maintain security, compliance, and governance standards as AI agents become more prevalent in business operations.",
"readTime": "2 mins",
"scrapedAt": "2026-09-01T17:09:28.876572+00:00",
"keyTakeaways": [
"AIR secured $50 million in funding to expand its AI agent management platform.",
"The platform discovers all AI agents operating within company infrastructure and systems.",
"It continuously vets agent capabilities, add-ons, and integrations to identify potential risks.",
"The system blocks unwanted behaviors and prevents unauthorized agent actions from executing.",
"The solution addresses enterprise concerns about AI agent proliferation and the need for centralized oversight and control."
]
}

For more information on the Actor's output, check the output schema.

DISCLAIMER: The developer of this Actor is not affiliated in anyway whatsoever with the Newscrusher.ai website.

Actor Input Options

  1. Maximum News Articles (maxResults) — The maximum number of articles to scrape in each Actor run.
  2. Sort (sort) — Specifies how to sort the articles feed (hot, latest, top-news, or top-research).
  3. Time Span (timeSpan) — The time span in which to look for articles. Options are today, yesterday, and this-week. If your sort is set to hot, only a timeSpan of today is supported. If you choose any other option in this case, it will be automatically considered today. If your sort is set to latest, the this-week time span is not supported and will be considered today. These limitations are NOT the Actor's fault. That's just how the website behaves.
  4. Include Summaries (includeSummaries) — Whether to include the article summaries in the Actor's output. Articles with their summaries included cost slightly higher than articles without summaries. This option is set to true by default. Be sure to disable it if you don't need summaries.
  5. Tags Section — This is a section for the tags to be enabled. Each tag is an option with the tag's name and can be set to true or false for enabling and disabling, respectively. You can enable multiple tags.

For more information on the Actor's input options, check the input schema.

How to Use the Actor?

To quickly get started with this Actor, follow the steps below:

  1. Sign in to Apify with your account or create a FREE account (no credit card required).
  2. Once you're signed in, get back to this Actor's page and click on the blue Try for free button.
  3. Once you're in the console, you'll see the input form of the Actor. Configure your input options as explained in the previous section and click on the green Save and start button at the bottom.
  4. Wait for the Actor to finish, then see your results in the output section. You can watch the Actor's progress during the run by going to the Log tab in the output section.
  5. Once the Actor finishes, you can export your results in many formats such as CSV, JSON, XLSX, and more. You can also schedule the Actor to run automatically in your desired interval. You can also set integrations such as automatically uploading results to google drive after each run.

❓ FAQs

What websites does this Actor scrape?

This Actor scrapes Newscrusher.ai, which aggregates AI news from a wide range of publishers, including TechCrunch, The Verge, Wall Street Journal, Quartz, Futurism, OpenAI, AWS ML Blog, and many more.

What data does the Actor return?

Each scraped article includes its headline, sub-headline, publisher, publication date, article URL, priority score, tags, scraping timestamp, and, when enabled, an AI-generated summary, estimated read time, and 5 key takeaways.

Are the article summaries AI-generated?

Yes. When Include Summaries (includeSummaries) is enabled, the Actor returns AI-generated summaries, article tags, estimated read time, and key takeaways. Disable this option to reduce the cost per article.

Can I filter the news I scrape?

Yes. You can choose the feed Sort (hot, latest, top-news, or top-research), select a Time Span (today, yesterday, or this-week), and enable specific tags to control which articles are included.

Why are some time span options unavailable?

Some combinations are limited by how Newscrusher.ai's website works. For example, hot only supports today, while latest does not support this-week. Unsupported combinations are automatically treated as the supported alternative.

Is this Actor affiliated with Newscrusher.ai?

No. This Actor is independently developed and is not affiliated with or endorsed by Newscrusher.ai.