# Tech Stack Detector (Wappalyzer & BuiltWith alternative) (`mambo_melon/tech-stack-detector`) Actor

Find out what any website is built with: CMS, shop platform, analytics, marketing and chat tools, hosting, email provider and SSL. Paste a list of domains, get one row per site. Pay per site checked.

- **URL**: https://apify.com/mambo\_melon/tech-stack-detector.md
- **Developed by:** [Joran Morgan](https://apify.com/mambo_melon) (community)
- **Categories:** Developer tools, Lead generation, SEO tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $5.00 / 1,000 website analyzeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Tech Stack Detector (Wappalyzer & BuiltWith alternative)

Give it a list of websites and it tells you what each one is built with: CMS, shop platform, analytics, marketing and chat tools, CDN, hosting, email provider and the SSL certificate. One row per website, ready for a spreadsheet or your CRM.

I use it for lead generation. A plumber still running a stock WordPress theme on shared hosting, with no booking widget and no chat, is a much warmer prospect for a web agency than a random name from Google Maps. Checking that by hand takes a few minutes per site. This goes through the whole list while you do something else.

It recognizes over 7,600 technologies. You pay per website that was actually analyzed. If a site doesn't load (dead domain, timeout, blocked), you still get a row explaining why, but it isn't charged.

### Example

Input:

```json
{ "urls": ["joesplumbingandheating.com", "ghost.org", "basecamp.com"] }
```

One of the rows you get back (shortened, real output):

```json
{
  "inputUrl": "https://joesplumbingandheating.com",
  "domain": "joesplumbingandheating.com",
  "statusCode": 200,
  "title": "Joe's Plumbing and Heating - Your home comfort starts here!",
  "techCount": 15,
  "techNames": ["Cloudflare", "WordPress", "MySQL", "GoDaddy", "PHP", "Yoast SEO",
                "WordPress Block Editor", "GoDaddy CoBlocks", "Twenty Twenty-Two", "..."],
  "technologies": [
    { "name": "WordPress", "version": null, "categories": ["CMS", "Blogs"], "confidence": 100, "website": "https://wordpress.org" },
    { "name": "Yoast SEO", "version": "28.5", "categories": ["SEO", "WordPress plugins"], "confidence": 100, "website": "https://yoast.com/wordpress/plugins/seo/" }
  ],
  "categories": { "CDN": ["Cloudflare"], "CMS": ["WordPress"], "Hosting": ["GoDaddy"] },
  "emailProvider": null,
  "hosting": "GoDaddy",
  "sslIssuer": "Google Trust Services",
  "sslExpiresAt": "Nov 12 10:38:42 2026 GMT",
  "sslValid": true,
  "sslError": null,
  "error": null,
  "checkedAt": "2026-09-29T09:05:51+00:00"
}
```

Every row has the same set of fields. If something couldn't be determined, it's `null` rather than missing, so exports to CSV/Excel don't shift columns.

A few fields worth knowing about:

- `emailProvider` comes from the domain's MX records (Google Workspace, Microsoft 365, Zoho, Proofpoint...). If the provider isn't one I recognize, you get the MX host itself.
- `hosting` is the most specific hosting/CDN signal found. Treat it as a hint; sites behind Cloudflare often hide the real host.
- `sslValid` / `sslError` tell you if the certificate is expired, mismatched or missing. Handy for spotting neglected sites.

### Input

| Field | What it does | Default |
|---|---|---|
| `urls` | Domains (`example.com`) or full URLs. Duplicates are dropped. | required |
| `maxItems` | Cap on websites per run | 1000 |
| `checkDns` | Look up MX/TXT/NS records (email provider, tools verified via DNS) | on |
| `maxConcurrency` | Sites checked in parallel | 20 |
| `requestTimeoutSecs` | How long to wait for a slow site | 20 |

### Things people use it for

Finding prospects by what they're missing (no analytics, no booking tool, outdated CMS), adding "runs Shopify / HubSpot / Salesforce" to a lead list before outreach, checking what competitors use for payments, ads or email, and measuring how common a tool is across a list of sites.

You can schedule it, call it from the API, or plug it into n8n, Make, Zapier or Google Sheets through the usual Apify integrations. It also works as a tool for AI agents via the Apify MCP server.

### FAQ

**Why not just use Wappalyzer or BuiltWith?**
They're good, but their APIs are subscriptions. If you check a few hundred sites a month, paying per site is a lot cheaper.

**Does it run a browser?**
No. It reads the HTML, response headers, cookies, script and meta tags, DOM selectors, DNS records and the TLS certificate. That catches almost every CMS, shop, analytics and marketing tool and keeps runs fast and cheap. The trade-off: a tool that only shows up after JavaScript runs in the browser can occasionally be missed.

**A site came back with an error.**
The `error` field says why: the domain doesn't resolve, the site timed out, or it answered 403/429 (some big sites block automated requests). Those rows aren't charged.

**How do I track changes over time?**
Run it on a schedule and compare `techNames` between runs, for example in Google Sheets.

### Notes

Only public pages are read. No logins, no personal data.

Detection rules come from [webappanalyzer](https://github.com/enthec/webappanalyzer) (GPL-3.0), the community-maintained continuation of the open-source Wappalyzer rules. I update them regularly.

If you're prospecting for a web agency, [Website Audit & Lead Scorer](https://apify.com/mambo_melon/website-audit-lead-scorer) uses the same detection and adds a lead score, the list of site problems and a first outreach line for each business.

Found a site it gets wrong? Open an issue with the URL and I'll take a look.

# Actor input Schema

## `urls` (type: `array`):

Websites to detect the tech stack of. Plain domains (example.com) or full URLs. Duplicates are removed.

## `maxItems` (type: `integer`):

Maximum number of websites to analyze in this run.

## `checkDns` (type: `boolean`):

Also read MX/TXT/NS records to detect email provider (Google Workspace, Microsoft 365...), DNS/CDN and marketing tools verified via DNS.

## `maxConcurrency` (type: `integer`):

How many websites to analyze in parallel.

## `requestTimeoutSecs` (type: `integer`):

Timeout for loading one website.

## Actor input object example

```json
{
  "urls": [
    "wordpress.org",
    "allbirds.com",
    "wix.com"
  ],
  "maxItems": 1000,
  "checkDns": true,
  "maxConcurrency": 20,
  "requestTimeoutSecs": 20
}
```

# Actor output Schema

## `overview` (type: `string`):

One row per website: detected technologies, email provider, hosting and SSL status.

## `detailed` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "urls": [
        "wordpress.org",
        "allbirds.com",
        "wix.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("mambo_melon/tech-stack-detector").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "urls": [
        "wordpress.org",
        "allbirds.com",
        "wix.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("mambo_melon/tech-stack-detector").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "urls": [
    "wordpress.org",
    "allbirds.com",
    "wix.com"
  ]
}' |
apify call mambo_melon/tech-stack-detector --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,mambo_melon/tech-stack-detector"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/1T6avinJYLZRsxXzb/builds/T8lQHsHBKbUGheZml/openapi.json
