Impressum Scraper – German, Austrian & Swiss Company Data avatar

Impressum Scraper – German, Austrian & Swiss Company Data

Pricing

from $1.50 / 1,000 imprint founds

Go to Apify Store
Impressum Scraper – German, Austrian & Swiss Company Data

Impressum Scraper – German, Austrian & Swiss Company Data

Extract company data from the Impressum (imprint / legal notice) of any DE, AT or CH website: company name, legal form, managing directors, HRB/HRA/FN register number and court, USt-IdNr/UID VAT ID, address, phone, fax and email. No Google search. Pay only when an imprint is found.

Pricing

from $1.50 / 1,000 imprint founds

Rating

0.0

(0)

Developer

Cemal Atakli

Cemal Atakli

Maintained by Community

Actor stats

0

Bookmarked

3

Total users

2

Monthly active users

6 days ago

Last modified

Categories

Share

Give it a list of websites from Germany, Austria or Switzerland. For each one, the Actor finds the Impressum (imprint / legal notice / Offenlegung) and returns the company data in one clean row: company name, legal form, managing directors (Geschäftsführer / Vorstand / Inhaber), register court and number (HRB, HRA, FN, CHE-UID), VAT ID (USt-IdNr., ATU…, CHE-… MWST), street, postcode, city, country, phone, fax, email and the person responsible for content. Each row also has a confidence score and the raw imprint text the fields were read from.

  • Your domains, no Google search. You bring the list (a CRM export, trade-fair exhibitors, a Google Maps scrape, a directory) and the Actor goes straight to each site. Nothing depends on search-engine scraping, so runs are fast, cheap and predictable.
  • Finds the imprint page itself. It reads the homepage links named Impressum, Imprint, Legal notice, Offenlegung, Anbieterkennzeichnung, Rechtliche Hinweise and similar (also in footers). If there is none, it tries common paths such as /impressum and /de/impressum. It follows "Legal" overview pages to the actual imprint and also handles one-page sites with the imprint on the homepage.
  • Built for DACH legal forms and registers. GmbH, UG (haftungsbeschränkt), AG, SE, KG, OHG, GbR, e.K., e.V., eG, PartG mbB, GmbH & Co. KG, SE & Co. KG, Austrian GmbH/Ges.m.b.H./OG/KG, Swiss AG/SA/GmbH/Sàrl. Register entries from Amtsgericht (HRB/HRA/VR/GnR/PR), Austrian Firmenbuch (FN 123456 a, Landesgericht/Handelsgericht) and Swiss UID (CHE-123.456.789) and Handelsregisteramt. For a GmbH & Co. KG, it returns the company's own HRA, not the general partner's HRB.
  • Honest output. Every row has confidence (0–1), confidenceLevel and fieldsFound, plus rawSnippet (the first 1,500 characters of the imprint block) so you can verify the data or post-process it with an LLM.
  • Pay only for hits. $1.50 per 1,000 websites where an imprint with a company name or a register/VAT number was found. Sites without an imprint, dead domains and blocked sites are free.
  • Polite. HTTP only (no browser), robots.txt respected (including Crawl-delay), a configurable delay between requests to the same host (1 s by default), sequential requests per site, and an honest User-Agent with a contact address.

Related Actors: Website Contact Finder for emails, phones and social profiles of any website worldwide, and Domain Checker – WHOIS/RDAP, DNS, SSL to check the domains themselves.

What can I use it for?

  • B2B lead enrichment in DACH: turn a list of company domains into legal name, managing director, address and register number for your CRM
  • KYC / KYB and supplier onboarding: check a supplier's legal entity, register number and VAT ID before you contract with them, then confirm them in the official register or VIES
  • Data cleaning: fix company names and legal forms in your CRM (e.g. "Muster" → "Muster Maschinenbau GmbH & Co. KG")
  • Market research: map which legal forms, cities and registers the companies in a niche use
  • Compliance monitoring: check that your own shops or your franchise partners publish a complete Impressum (register, VAT ID, responsible person)
  • AI agents: give an agent a "who operates this German website?" tool (see below)

Input

FieldDescription
urlsDomains or URLs (muster.de, https://www.firma.at/), or the imprint URL directly
bulkTextPaste a big list or a CSV export (the first domain-like cell per row is used)
sourceFileUrlPublic TXT/CSV file with domains (e.g. a Google Sheets "publish to web" CSV link)
onlyFoundLeave out rows for sites where no imprint data was found (default: off)
includeRawSnippetAdd the raw imprint text (default: on)
maxSites, maxPagesPerSite (default 6), respectRobotsTxt (default on), minDelayPerHostSecs (default 1), maxConcurrency (default 10), requestTimeoutSecs, proxyConfigurationAdvanced
{
"urls": ["heise.de", "zotter.at", "kaffeemacher.ch"],
"onlyFound": false,
"includeRawSnippet": true
}

Output

One row per website. The Output tab has three tables: Company data, Contacts & people and Register & tax IDs. Here is a real row (rawSnippet shortened):

{
"domain": "heise.de",
"website": "https://www.heise.de/",
"imprintUrl": "https://www.heise.de/impressum.html",
"imprintFound": true,
"companyName": "Heise Medien GmbH & Co. KG",
"legalForm": "GmbH & Co. KG",
"managingDirectors": ["Ansgar Heise", "Beate Gerold"],
"representatives": [{"name": "Ansgar Heise", "role": "Geschäftsführer"}, {"name": "Beate Gerold", "role": "Geschäftsführer"}],
"registerCourt": "Amtsgericht Hannover",
"registerType": "HRA",
"registerNumber": "26709",
"registerId": "HRA 26709",
"vatId": "DE813501887",
"street": "Karl-Wiechert-Allee 10",
"postalCode": "30625",
"city": "Hannover",
"country": "DE",
"phone": "+49 511 53520",
"fax": "+49 511 5352129",
"email": "webmaster@heise.de",
"responsiblePerson": "Dr. Volker Zota",
"confidence": 1.0,
"confidenceLevel": "high",
"rawSnippet": "Impressum\nVerantwortlich für dieses Angebot:\nHeise Medien GmbH & Co. KG\nKarl-Wiechert-Allee 10\n30625 Hannover…",
"error": null
}

An Austrian row has registerType: "FN", for example "registerId": "FN 220619s" with "registerCourt": "Landesgericht für ZRS Graz" and "vatId": "ATU53816900". A Swiss row has "registerId": "CHE-256.970.360", "swissUid" and "vatId": "CHE-256.970.360 MWST". Other fields: taxNumber (Steuernummer), emails (all emails in the imprint), companyNameSource, imprintSource, pagesFetched, robotsTxt, blocked, durationMs. See SAMPLE_OUTPUT.json for 13 real rows, including a sole trader, an e.V., a robots.txt-disallowed shop and a blocked site.

Accuracy (local test, October 2026). We ran 52 real DE/AT/CH websites, from sole traders and blogs to Kärcher, Tchibo, BILLA and Ricola. 44 returned imprint data. Of the other 8, 6 were bot-protected (HTTP 403/406 or challenge pages), 1 Shopify shop disallows its legal-notice page in robots.txt and 1 site was too slow. We checked the fields by eye against the live imprint for 27 sites:

  • Company name: 27/27 correct
  • Register number: 23/23 correct
  • VAT ID: 23/23 correct
  • Managing directors: 20/21 complete (one named only in a parenthesis was missed)
  • Address: 26/27 complete (one street came back only partly)
  • Phone, fax and email: all correct where present

Swiss imprints often omit the UID. In that case you get name, address and contacts at confidence 0.5.

Pricing

Pay per event:

EventPrice
Imprint found (company name or register/VAT number extracted)$0.0015 ($1.50 per 1,000)
Actor start$0.00005
No imprint, dead domain, blocked site, robots-disallowed pagefree

How that compares with other imprint scrapers in the Apify Store (listed prices, October 2026):

ActorListed priceNotes
winningsolutions/german-imprint-scraper$5 / 1,000finds sites through Google search
dominic-quaiser imprint scraper$1.20 / 1,000
This Actor$1.50 / 1,000 hits; misses are freeyour own domain list, DE + AT + CH registers, confidence + raw text

Set a maximum cost per run in the run options. The Actor saves as many hits as fit and then stops cleanly.

GDPR and responsible use

An Impressum is published because the law requires it (§ 5 DDG in Germany, § 5 ECG and § 25 MedienG in Austria, Art. 3 UWG in Switzerland). It is meant to let people identify and contact the business. It still often contains personal data, such as managing directors' names, sole traders' home addresses and personal emails.

  • This Actor is intended for B2B use: identifying the company behind a website, KYB/supplier checks, and enriching business records.
  • You are the controller of the data you collect and are responsible for a lawful basis (usually legitimate interest, Art. 6(1)(f) GDPR), data minimisation, retention limits, and informing data subjects where required (Art. 14 GDPR).
  • An Impressum is not consent to marketing. Cold emails or calls to businesses are restricted in Germany (§ 7 UWG) and Austria (§ 174 TKG). Check the rules for your channel before you use the contacts for outreach.
  • Use includeRawSnippet: false and drop the fields you do not need. Verify register numbers in the official registers (Handelsregister, Firmenbuch, Zefix) and VAT IDs in VIES before you rely on them.
  • The Actor reads only publicly accessible pages, respects robots.txt, and does not log in or bypass bot protection.

FAQ

How does it find the imprint? It reads the homepage links whose text or URL says Impressum, Imprint, Legal notice, Offenlegung, Rechtliche Hinweise and so on. If none is found, it tries /impressum, /de/impressum, /imprint, /kontakt and similar paths. If the first page has no company data (for example a "Legal" overview page), it follows the imprint links on that page. It fetches at most maxPagesPerSite pages per site (default 6, usually 2 are enough).

Why is a site marked "Blocked by bot protection"? Some large shops (Cloudflare, Akamai and similar) block every non-browser client. This Actor does not use browsers or try to bypass protection. Those rows are free.

Why is the imprint of a Shopify shop "disallowed by robots.txt"? Shopify's default robots.txt disallows /policies/, which is where Shopify shops keep their legal notice. With respectRobotsTxt on (the default and our recommendation) the page is not fetched, and you pay nothing for that row.

What is confidence? A 0–1 score built from what was found: company name with a legal form, register number and court, VAT ID, a full address, managing directors and contact data. 0.7 or higher is high. Sole traders without a register usually score 0.3–0.5. The score is higher there when name and address are present.

Does it work for other countries? It is tuned for the German-speaking imprint format (DE, AT, CH, LI). English-language legal notices of DACH companies ("Registration court", "VAT No.", "Managing directors") work too. For other countries use Website Contact Finder.

Can I get CSV or Excel? Yes. Every Apify dataset can be exported to CSV, Excel, JSON or XML, or sent to Google Sheets with an integration. Array fields such as managingDirectors become joined columns.

Use with AI agents / Apify MCP

The input is just a domain and the output is structured company data at $0.0015 per hit, which makes this Actor a good agent tool. Connect it through the Apify MCP server (https://mcp.apify.com?actors=gazidev/imprint-scraper) and Claude, ChatGPT, Cursor and other MCP clients can answer prompts such as "Who is the managing director of the company behind muster-shop.de, and what is its HRB number?" or "Enrich these 200 Austrian domains with FN number and UID". Over the API: POST https://api.apify.com/v2/acts/gazidev~imprint-scraper/run-sync-get-dataset-items with the input JSON.

Website toolkit by gazidev

Other low-cost, HTTP-only Actors for working with lists of websites:

Categories

Lead generation · Business · Automation