Workplace Program Detector ERG DEI Benefits Parental Leave avatar

Workplace Program Detector ERG DEI Benefits Parental Leave

Pricing

from $6.80 / 1,000 domain analyzeds

Go to Apify Store
Workplace Program Detector ERG DEI Benefits Parental Leave

Workplace Program Detector ERG DEI Benefits Parental Leave

Detects which people programs a company publishes: employee resource groups, DEI, wellbeing and mental health, learning and tuition support, parental and caregiver leave, volunteering. Reads careers, culture, benefits and ESG pages plus live job postings. One flat row per domain for Clay.

Pricing

from $6.80 / 1,000 domain analyzeds

Rating

0.0

(0)

Developer

Mamba Labs

Mamba Labs

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

4 hours ago

Last modified

Share

๐Ÿ”Ž What can Workplace Program Detector do?

Give it a company domain and it returns one flat row naming the people programs that company publishes: employee resource groups, DEI, wellbeing and mental health, learning and tuition support, parental and caregiver policy, and volunteering and giving.

Every row also reports how much of the company the actor actually read, so you can tell a real "no program" apart from a page it could not open.

๐Ÿ“ฆ What you getโš™๏ธ Features and integrations
๐Ÿ… Six program families, one boolean each
๐Ÿ”— Evidence URL and phrase per family, so you can check the finding
๐Ÿ“Š Coverage block, what was reached, blocked or thin
๐Ÿงพ 34 flat fields, snake_case, one row per domain
๐ŸŒ Two retrieval paths, website pages and live job postings
๐Ÿค– Auto board discovery for Greenhouse, Lever and Ashby
๐ŸงŠ 14 day cache, with a skipCache override
โฌ‡๏ธ Export to JSON, CSV, Excel, HTML or XML

Bought by teams selling into HR, People Ops, benefits and learning and development, and by employer brand and DEI consultancies building prospect lists from what companies publish.

๐Ÿšซ This is not an employer review or ratings source. It does not score a company, rate its culture, or tell you whether a program works. It records what the company published on its own website and job board, and nothing else.

๐Ÿ’ก Why use Workplace Program Detector?

If you sellRead these fields
ERG and community platformshas_erg, erg_evidence_url, erg_evidence_phrase
DEI programs and consultinghas_dei_program, dei_evidence_url
Mental health and wellbeing benefitshas_wellbeing_program, wellbeing_evidence_url
Learning platforms and tuition benefitshas_learning_program, learning_evidence_url
Family and caregiver benefitshas_parental_policy, parental_evidence_url
Volunteering and giving platformshas_volunteering_program, volunteering_evidence_url
Anything, as a disqualifierprogram_count, coverage, pages_reached

๐Ÿงญ Two retrieval paths, because one path misses half the answer

The actor reads company web pages and live job postings, and it ships both because they find different things.

Measured during the build: across five real job boards covering 1,764 open roles, employee resource groups appeared in zero job bodies. Across seven company websites, employee resource groups appeared on two. Benefits language runs the other way, and shows up in job bodies that marketing pages leave out.

signal_source on every row tells you which path produced the finding.

๐Ÿ“‹ What data can Workplace Program Detector extract?

34 fields per domain. The ones buyers use:

FieldWhat it holds
has_ergCompany publishes employee resource groups
has_dei_programPublished DEI program, excluding legal boilerplate
has_wellbeing_programWellbeing or mental health program
has_learning_programLearning, development or tuition support
has_parental_policyParental or caregiver policy
has_volunteering_programVolunteering or giving program
program_countHow many of the six were found
*_evidence_urlThe page the finding came from
*_evidence_phraseThe wording that matched
signal_sourceWhich path fired, website or job postings
coverageHow much of the company was read
pages_attempted, pages_reached, pages_reached_listWhat was tried and what opened
pages_blocked, pages_thinRefused pages, and pages with almost no readable text
ats_provider, ats_slug_used, jobs_scannedWhich job board was read, and how many jobs
render_modeclient_rendered when the page needs a browser to show text
fetch_status, fetch_errorok, blocked, or the error that stopped it

โš ๏ธ false and null are not the same thing. false means the pages were read and the language was not there. null means not enough was read to say. They are never collapsed into each other. A company that returns 403 on every path comes back with all six fields null and fetch_status: "blocked", not six confident falses. Read coverage and pages_reached before you trust any false.

๐Ÿ› ๏ธ How to find a company's employee programs

  1. Open the Input tab and put a bare domain in domain, for example hubspot.com.
  2. Leave scan_web_pages and scan_job_postings on true so both paths run.
  3. Click Start.
  4. Read program_count and the six has_* booleans for the answer.
  5. Read coverage and pages_reached before you act on any false.
  6. Export from the Output tab, or pull the row through the API.

๐Ÿงช Using it in Clay

Add it as an Apify enrichment, map your domain column to domain, and the 34 fields land as one flat row with no reshaping. Every input is accepted as a string, which is what Clay sends.

Gate the run on your ICP column so you only spend the event on accounts you would actually work.

โšก Skipping board discovery

If you already know the company's job board slug, put it in ats_slug and the actor reads that board directly instead of discovering it. That is the fastest path when you are running a list you have already resolved.

๐Ÿ’ต How much does it cost to detect workplace programs?

One domain-analyzed event per domain.

PlanPrice per domain
Free$0.008
Bronze$0.0076
Silver$0.0072
Gold$0.0068

There is also an Actor start event at $0.00005, charged once per run per GB of memory.

๐Ÿ’ณ A domain that turns out to be unreachable is still billed. The actor did the work of attempting it, and a null row that tells you the site blocked us is a real answer. What is never billed is a run that does not start.

โŒจ๏ธ Input

Everything is on the Input tab. The options worth explaining:

FieldTypeDefaultWhat it does
domainstringrequiredBare domain, for example hubspot.com. Protocol and path are stripped.
ats_slugstringemptySkip board discovery and read this Greenhouse, Lever or Ashby slug directly.
scan_job_postingsbooleantrueRead the company's live job bodies.
scan_web_pagesbooleantrueProbe the careers, culture, benefits, DEI and ESG paths.
max_pagesstring14Clamped to 1 to 25. Sent as a string for Clay.
skipCachestringfalsetrue forces a fresh crawl past the 14 day cache.

๐Ÿ“ค Output

One row per domain, exportable as JSON, CSV, Excel, HTML or XML. 34 flat snake_case fields, with null rather than a missing key.

{
"domain": "hubspot.com",
"has_erg": true,
"erg_evidence_url": "https://www.hubspot.com/careers/culture",
"erg_evidence_phrase": "employee resource groups",
"has_dei_program": true,
"has_wellbeing_program": true,
"has_learning_program": true,
"has_parental_policy": true,
"has_volunteering_program": false,
"program_count": 5,
"signal_source": "website",
"coverage": "high",
"pages_attempted": 14,
"pages_reached": 11,
"pages_blocked": 0,
"pages_thin": 1,
"ats_provider": "greenhouse",
"jobs_scanned": 132,
"render_mode": "server_rendered",
"fetch_status": "ok",
"fetch_error": null
}

๐Ÿ’ก Tips

  • Run it monthly on the same list and diff two rows. That turns a snapshot into a trend.
  • A documented absence is a pitch. has_wellbeing_program: false with coverage: "high" is a company whose competitors publish a program and it does not.
  • Use program_count as a cheap sort. Companies publishing five or six families are already investing in this area and are usually the warmer conversation.
  • Set ats_slug when you have it. It removes the discovery step.

โš ๏ธ Known limits

A false is only as good as the pages that opened. On a company where one page was reached, a false means very little. Check pages_reached and coverage on every row.

JavaScript-only careers sites cannot be read. Some large companies serve a careers page that is an empty shell until a browser runs it. Measured during the build: one site's careers page held 20 characters of readable text, and another held 1,141 characters inside 84,600 bytes of markup. Those rows say render_mode: client_rendered with pages_thin above zero. Treat every false on them as unknown.

Some sites refuse outright. One of the seven domains in the build sample returned HTTP 403 on nine of fourteen paths. Those rows come back fetch_status: "blocked" with all six program fields null.

Some sites ask not to be read. robots.txt is read and honored on every domain before probing. A site that disallows crawling gets pages_attempted: 0 and null program fields. That is the site's decision recorded correctly, not a company without programs.

Employee resource groups are almost never named in job postings. If a company blocks its website, has_erg stays null even when the other families resolve from job postings.

Internal branding defeats keyword matching. Companies name their groups whatever they like. One large company in the build sample calls them Equality Groups and was invisible until that phrase was added. Expect misses on unusual internal naming.

Small companies often have no reachable website. One test target has no working HTTPS certificate and returns a network error on every path. That is a none coverage row, not a company without programs.

This is a snapshot, not a trend. The actor does not say whether a program is growing or being wound down. The archival sources that would answer that take between 2 and 18 seconds per query, far more than this actor's whole runtime. Run it monthly and diff instead.

Legal boilerplate is excluded on purpose. "Equal opportunity employer" appears in almost every US job posting and is not counted as a DEI program.

The actor records that a company publishes employee resource groups, and deliberately does not record which groups. The ERG evidence phrase is capped tight and dropped entirely if what survives names a protected characteristic. When that happens has_erg stays true and the evidence URL stays populated, so you keep the finding and its source and lose only the sentence. The same scrub runs on the DEI phrase.

โ“ FAQ

Why is every program field null?

The site blocked the actor or could not be reached. Check fetch_status and pages_blocked. That row is a coverage problem, not a company without programs.

Does it use a proxy?

No. Every path runs direct with full browser headers. The design answer to a block is an honest blocked status rather than a workaround.

How fresh is the data?

Results are cached for 14 days per domain and input combination, because careers and benefits pages change quarterly at best. Pass skipCache: "true" to force a fresh crawl.

Can I run a list instead of one domain?

Yes. Run it per row from Clay, or drive it through the Apify API and collect the dataset.

Does it tell me the names of a company's ERGs?

No, by design. It records that a company publishes employee resource groups and drops any evidence phrase that names a protected characteristic.

๐Ÿงฉ Want other GTM data?

Mamba Labs builds custom actors for B2B go-to-market teams. The public versions of that work live here on the Store, so our users get the same tooling we build under contract.

๐Ÿง‘โ€๐Ÿ’ผ GTM Hiring Signal Scraper๐Ÿงฑ Tech Stack Detector
๐Ÿ“ก B2B Buying Signals Aggregator๐Ÿ”‘ Job Board Keyword Scanner
๐Ÿ”— Domain to LinkedIn URL Resolver๐ŸŽฏ ICP Fit Scorer
๐Ÿ“‹ Job Posting Monitor๐Ÿ“ฌ Domain Deliverability Checker
๐Ÿข Company Firmographic Enricher๐ŸŒ Company Social Presence Mapper
๐Ÿชช Company Identity Resolver๐Ÿ’ฐ Funding and Press Signal Scanner
๐Ÿ”„ Company Change-Event Feed๐Ÿ‘ค People Finder and Email Verifier
๐Ÿš€ Prospect Engine๐Ÿค– AI Tooling Detector
๐Ÿ“ฎ Outbound Stack Detector๐Ÿ“ Publishing Frequency Tracker
โœ‰๏ธ Work Email Waterfall Finderโฉ Sequencer Lead Push
๐Ÿ‘ฅ Team Page People Extractor๐Ÿงญ Company Discovery List Builder

Every actor in the suite takes a domain or a company and returns one flat row, so they stack in the same Clay table without reshaping anything.

๐Ÿ› ๏ธ Need something custom built for you or your team? Tell us what you are trying to find and we will build it. Talk to Mamba Labs.

๐Ÿ†˜ Support

Something wrong, or a company the actor reads incorrectly? Open an issue on the Issues tab with the domain and the row, and we will look at it.

โ„น๏ธ Sourcing and legal. Every field comes from pages the company publishes itself, read directly, with robots.txt honored on every domain. The row describes what a company published on its own website and job board. It is not an assessment of the company, its workforce, or how well any program works. You are responsible for how you use the output, including under applicable employment and data protection law.

Built by Mamba Labs.