# DER STANDARD News Scraper (`rainminer/derstandard-at-scraper`) Actor

Extract news articles from DER STANDARD Austria — headlines, authors, dates, categories, excerpts, images, and full article text. Export structured JSON/CSV via Apify.

- **URL**: https://apify.com/rainminer/derstandard-at-scraper.md
- **Developed by:** [rainminer](https://apify.com/rainminer) (community)
- **Categories:** News, Lead generation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.49 / 1,000 articles

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

![DER STANDARD](https://www.derstandard.at/favicon.ico)

### What does DER STANDARD News Scraper do?

**DER STANDARD News Scraper** extracts **public news articles** from [derstandard.at](https://www.derstandard.at/) — Austria's leading online newspaper. Paste a section URL (for example `https://www.derstandard.at/international`), an RSS feed, a date archive, or a single story URL and export structured content — **headline, author, date, category, excerpt, image, and full article text** — without writing any code.

Run it on the **Apify platform** to schedule crawls, export JSON/CSV/Excel, rotate proxies, monitor failures, and integrate results via API.

### What can this DER STANDARD scraper do?

- **Section and RSS discovery** — Start from International, Wirtschaft, Sport, Inland, or any `/rss/{section}` feed.
- **Full article text** — Visits each story page and exports body text, kicker, author, and hero image.
- **Headline-only mode** — Set `includeFullContent` to `false` for fast RSS exports without article pages.
- **Date archives** — Scrape a specific day's coverage via `/section/YYYY/M/D` URLs with automatic older-day pagination.
- **Structured output** — Title, excerpt, content, publish date, category, image URL, and canonical story URL.
- **No login required** — Public Austrian news only; lightweight HTTP crawler, no browser needed by default.

### Why scrape derstandard.at?

Use cases for **DER STANDARD** and **Austrian news** data:

1. **Austrian media monitoring** — Track breaking politics, business, and culture headlines in one feed.
2. **PR and communications** — Watch how topics evolve across DER STANDARD sections.
3. **Research and journalism** — Collect article metadata and full text for trend analysis.
4. **Content aggregation** — Build newsletters or dashboards from International, Wirtschaft, or Sport.
5. **Competitive intelligence** — Monitor Austrian market and policy stories by section.
6. **NLP and RAG datasets** — Export full article bodies for summarization or embeddings.
7. **Historical news archives** — Walk dated archive pages for day-by-day coverage.
8. **Cross-border news tracking** — Follow international stories relevant to Austria and the EU.
9. **Editorial research** — Study kickers, bylines, and lead paragraphs across topics.
10. **Alerting workflows** — Schedule runs and push new headlines to Slack, email, or webhooks.

Related Actor: [DER STANDARD Jobs Scraper](https://apify.com/rainminer/derstandard-jobs-at-scraper) for `jobs.derstandard.at` listings.

### How to use DER STANDARD News Scraper

1. Open the Actor in Apify Console.
2. Add one or more **derstandard.at section, RSS, archive, or story URLs** under **Start URLs**.
3. Set **Maximum items** (articles per start URL).
4. Run the Actor and download the results from the **Dataset** tab.

### Input

| Field | Description |
| --- | --- |
| `startUrls` | Section, RSS, date archive, or `/story/` article URLs on `www.derstandard.at`. |
| `maxItems` | Maximum articles per start URL (default `5`). |
| `includeFullContent` | When `true`, visits each story for full body text (default). |
| `proxyConfiguration` | Optional proxy; **no proxy by default**. |

Supported start URLs:

- Section pages: `https://www.derstandard.at/international`, `https://www.derstandard.at/wirtschaft`
- RSS feeds: `https://www.derstandard.at/rss/international`
- Date archives: `https://www.derstandard.at/international/2026/9/23`
- Article pages: `https://www.derstandard.at/story/{id}/{slug}`

### Output

Each result is a structured article object. You can download the dataset in **JSON, CSV, Excel, or HTML**.

```json
{
  "id": "3000000340940",
  "title": "Israels „Kriegsverbrechen“, Trump und Irankrieg: Hitzige Debatten bei UN-Generaldebatte",
  "excerpt": "Israel weist Van der Bellens deutliche Kritik scharf zurück...",
  "content": "Bei einer Pressekonferenz mit Bundeskanzler Christian Stocker...",
  "publishedAt": "2026-09-23T08:58:10.085Z",
  "publishedAtText": "23. September 2026, 10:58",
  "author": "Flora Mory",
  "kicker": "New York",
  "category": "international",
  "imageUrl": "https://i.ds.at/.../image.jpg",
  "url": "https://www.derstandard.at/story/3000000340940/...",
  "searchUrl": "https://www.derstandard.at/international"
}
```

### What data can you extract from derstandard.at?

| Field | Description |
| --- | --- |
| id | DER STANDARD story ID |
| title | Article headline |
| excerpt | Lead paragraph / subtitle |
| content | Full article body text |
| publishedAt | ISO publication timestamp when available |
| publishedAtText | Human-readable date as shown on page |
| author | Byline / author name |
| kicker | Story kicker label (e.g. location or topic tag) |
| category | Section name derived from the start URL |
| imageUrl | Primary hero image URL |
| url | Public article URL |
| searchUrl | Source section or archive URL for articles found via search |

### How much does it cost to scrape derstandard.at?

This Actor uses **pay-per-event** billing on the **Article** dataset item. Platform usage is included in the event price for typical HTTP runs without proxy.

- **FREE tier:** about **$1.99 per 1,000 articles**
- **GOLD tier:** about **$1.49 per 1,000 articles**
- **Actor start:** $0.00005 per run

A test run with 5 articles per section usually finishes in under a minute on the Apify free tier. Larger crawls scale with article count and archive depth.

### Tips and advanced options

- Use section URLs like `https://www.derstandard.at/international` for the latest headlines via RSS.
- Set **Include full article text** to `false` for faster headline-only exports from RSS.
- Use date archive URLs to scrape a specific day's coverage.
- Increase **maxItems** to walk older archive pages when you need more than the latest RSS batch.

### FAQ, disclaimers, and support

**Is scraping derstandard.at legal?**\
This Actor collects only **public** news content shown on derstandard.at without login. You are responsible for complying with applicable laws when processing any data.

**Which URLs are supported?**\
Section pages, RSS feeds, dated archive pages under a section, and individual `/story/` article URLs on `www.derstandard.at`.

**Need help?**\
Open an issue on the Actor page or contact the developer through Apify Console.

### Image Credit

Image credit: [derstandard.at](https://www.derstandard.at/)

# Actor input Schema

## `startUrls` (type: `array`):

derstandard.at section URLs (e.g. https://www.derstandard.at/international), RSS feeds, date archives, or individual story URLs.

## `maxItems` (type: `integer`):

Maximum number of articles to return per start URL.

## `includeFullContent` (type: `boolean`):

When enabled, visits each story page and extracts the full article body. Disable for faster RSS headline-only exports.

## `proxyConfiguration` (type: `object`):

Proxy settings. The scraper works without a proxy in most regions.

## Actor input object example

```json
{
  "startUrls": [
    {
      "url": "https://www.derstandard.at/international"
    }
  ],
  "maxItems": 5,
  "includeFullContent": true,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# Actor output Schema

## `overview` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        {
            "url": "https://www.derstandard.at/international"
        }
    ],
    "proxyConfiguration": {
        "useApifyProxy": false
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("rainminer/derstandard-at-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": [{ "url": "https://www.derstandard.at/international" }],
    "proxyConfiguration": { "useApifyProxy": False },
}

# Run the Actor and wait for it to finish
run = client.actor("rainminer/derstandard-at-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    {
      "url": "https://www.derstandard.at/international"
    }
  ],
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}' |
apify call rainminer/derstandard-at-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,rainminer/derstandard-at-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/4Nyxb0dpV1njydgTE/builds/QjmW79grCjCLfBYrN/openapi.json
