IRS Tax-Exempt Organizations Database Exporter
Pricing
from $1.50 / 1,000 results
IRS Tax-Exempt Organizations Database Exporter
Export nearly 2 million US tax-exempt organizations from the official IRS EO BMF. Get EIN, name, address, NTEE, filing status, assets, income and revenue by state. No API key required.
Pricing
from $1.50 / 1,000 results
Rating
0.0
(0)
Developer
Logiover
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
5 days ago
Last modified
Categories
Share
Export nearly 2 million US tax-exempt organizations from the official IRS EO BMF. Get EIN, name, address, NTEE, filing status, assets, income and revenue by state. No API key required.
What does the IRS Tax Scraper do?
This Actor turns IRS Tax into a structured dataset you can actually work with. You give it a search, it walks the result pages, and it writes one clean row per record into your dataset โ ready to export as JSON, CSV or Excel, or to pull straight from the Apify API.
It runs on plain HTTP with no browser and no login, which keeps it fast and cheap, and it works the same on a free account as on a paid one. Every field documented below came from a real verification run of this Actor, so what you read here is what you get.
The current build returns 31 populated fields per record, of which 28 are present on virtually every row.
Who is it for?
- Analysts who need IRS Tax data as a spreadsheet instead of a browser tab.
- Operations and research teams tracking how IRS Tax listings change over time.
- Developers wiring a live data feed into an internal tool or database.
- Growth and sales teams building prospect and market lists from public data.
- AI teams assembling structured training or grounding data.
Use cases
- Export a full IRS Tax search into a spreadsheet for analysis.
- Track how prices, volumes or availability move week over week by scheduling the run.
- Feed a dashboard or internal database with a repeatable, structured source.
- Build a market map by running several searches in one job.
- Enrich an existing list by matching on the identifiers in the output.
Why use this IRS Tax Scraper?
- ๐ No API key and no login โ nothing to authenticate, nothing to expire.
- ๐ฆ 31 verified fields โ every column below was confirmed against live output.
- ๐ Real pagination โ it walks the result set instead of returning the first screen.
- ๐ฏ Bounded runs โ caps in the input stop the job exactly where you want, so the bill is predictable.
- ๐ธ Tiered pay-per-result โ you pay per row, and higher Apify plans pay less per row.
- ๐ Export anywhere โ JSON, CSV, Excel, HTML, the API, or any Apify integration.
What data can you extract?
One row per record. These columns come from a live run, with the share of rows that carried a value:
| Field | Type | Typically filled | Description |
|---|---|---|---|
accountingPeriod | number | 100% | Accounting period |
activityCodes | string | 100% | Activity codes |
affiliationCode | string | 100% | Affiliation code |
assetCode | string | 100% | Asset code |
assets | number | 100% | Assets |
city | string | 100% | City |
classificationCodes | string | 100% | Classification codes |
deductibilityCode | string | 100% | Deductibility code |
ein | string | 100% | Ein |
filingRequirementCode | string | 100% | Filing requirement code |
foundationCode | string | 100% | Foundation code |
groupExemptionNumber | number | 100% | Group exemption number |
income | number | 100% | Income |
incomeCode | string | 100% | Income code |
irsSearchUrl | string | 100% | Irs search url |
name | string | 100% | Name |
nteeCode | string | 100% | Ntee code |
organizationCode | string | 100% | Organization code |
postalCode | string | 100% | Postal code |
privateFoundationFilingCode | string | 100% | Private foundation filing code |
rulingDate | string | 100% | Ruling date |
sourceFileUrl | string | 100% | Source file url |
sourceUrl | string | 100% | Source url |
state | string | 100% | State |
statusCode | string | 100% | Status code |
street | string | 100% | Street |
subsectionCode | string | 100% | Subsection code |
taxPeriod | string | 100% | Tax period |
careOfName | string | 72% | Care of name |
revenue | number | 87% | Revenue |
sortName | string | 5% | Sort name |
Fields below 90% are genuinely optional at the source โ IRS Tax does not publish them for every record. An empty value means the source did not show it, not that extraction failed.
Output example
{"accountingPeriod": "12","activityCodes": "000000000","affiliationCode": "3","assetCode": "6","assets": 1124006,"city": "GARRISON","classificationCodes": "1000","deductibilityCode": "1","ein": "010597067","filingRequirementCode": "01","foundationCode": "16","groupExemptionNumber": "0000","income": 4070948,"incomeCode": "6","irsSearchUrl": "https://apps.irs.gov/app/eos/details/?ein=010597067&name=&city=&state=&countryAbbr=US&type=returnsSearch","name": "GARRISON INSTITUTE","nteeCode": "B99","organizationCode": "1","postalCode": "10524-0532","privateFoundationFilingCode": "0","rulingDate": "200206","sourceFileUrl": "https://www.irs.gov/pub/irs-soi/eo_ny.csv"}
How to use
- Open the Actor and fill in the search fields โ the defaults below are a working example.
- Set the result cap so the run stops where you want it.
- Click Start, then export from the Output tab or pull the dataset through the API.
{"stateCodes": ["NY","NJ"],"nteePrefix": "B","minAssets": 1000000,"maxResults": 10000}
Input parameters
| Parameter | Type | Default | Description |
|---|---|---|---|
stateCodes | array | ["CA"] | Two-letter codes. |
nameContains | string | โ | Optional case-insensitive organization name fragment. |
city | string | โ | Optional exact city name, matched case-insensitively. |
nteePrefix | string | โ | Optional nonprofit activity classification prefix, such as B for education or E for health. |
subsection | string | โ | Optional numeric subsection code, for example 03 for 501(c)(3). |
minAssets | number | โ | Optional minimum reported asset amount. |
minRevenue | number | โ | Optional minimum reported revenue amount. |
maxResults | integer | 1000 | Maximum matching organizations to save; supports bulk exports up to 500,000 items per run. |
Tips for best results
- Broad searches return the deepest result sets; very narrow ones run out after a page or two.
- Use the result cap rather than the page cap when you want a predictable bill.
- Schedule the same input daily or weekly to build a history instead of a one-off snapshot.
- Lower the concurrency if you see retries โ the source rate-limits aggressive crawling.
- Run several searches in one job instead of one search with a huge page count.
- Empty optional fields are normal; filter on the columns that matter to you after export.
- Deduplicate on the identifier column if you merge several runs together.
- Keep the proxy setting on the default unless the source blocks your region.
Integrations
Push results straight into Google Sheets, Slack, Zapier, Make, Airtable or any Webhook. You can also schedule the Actor and have each run append to the same dataset, which is how you turn a single export into a time series.
API usage
Replace <YOUR_TOKEN> with your Apify API token.
cURL
curl -X POST "https://api.apify.com/v2/acts/logiover~irs-tax-exempt-organizations-scraper/run-sync-get-dataset-items?token=<YOUR_TOKEN>" \-H "Content-Type: application/json" \-d '{"stateCodes": ["NY", "NJ"], "nteePrefix": "B", "minAssets": 1000000, "maxResults": 10000}'
Python
from apify_client import ApifyClientclient = ApifyClient('<YOUR_TOKEN>')run = client.actor('logiover/irs-tax-exempt-organizations-scraper').call(run_input={"stateCodes": ["NY", "NJ"], "nteePrefix": "B", "minAssets": 1000000, "maxResults": 10000})for item in client.dataset(run['defaultDatasetId']).iterate_items():print(item)
Node.js
import { ApifyClient } from 'apify-client';const client = new ApifyClient({ token: '<YOUR_TOKEN>' });const run = await client.actor('logiover/irs-tax-exempt-organizations-scraper').call({"stateCodes": ["NY", "NJ"], "nteePrefix": "B", "minAssets": 1000000, "maxResults": 10000});const { items } = await client.dataset(run.defaultDatasetId).listItems();console.log(items);
Use with AI agents (MCP)
This Actor is reachable through the Apify MCP server, so an AI agent can call it as a tool. Point the agent at https://mcp.apify.com and ask it something like "pull the first 200 results from IRS Tax and summarise them" โ it receives the same structured rows you would get from the UI.
FAQ
Do I need a IRS Tax account or API key?
No. The Actor reads publicly available pages only. There is nothing to authenticate and no credentials to rotate.
Does it work on the free plan?
Yes. Every feature is available on a free Apify account โ nothing is gated behind a paid tier. Pricing is per result, and higher plans simply pay less per row.
How many results can I get in one run?
As many as the search exposes. Raise the page and result caps together; the run stops at whichever limit it reaches first.
Why did I get fewer rows than I asked for?
The search ran out of records. That is normal for narrow queries โ broaden the search or add more searches to one run.
Why are some fields empty?
IRS Tax does not publish every attribute for every record. The table above shows how often each field is filled in practice.
What export formats are supported?
JSON, CSV, Excel, HTML and RSS from the Output tab, plus the Apify API and any integration you connect.
How fast is it?
It is pure HTTP with no browser, so a page of results typically takes a second or two.
Can I schedule it?
Yes. Use the Apify scheduler to run it on any interval and append each run to the same dataset.
Does it use a proxy?
It routes through Apify Proxy by default. You can switch groups or supply your own proxies in the input.
Is the output schema stable?
Yes. Field names and types are fixed, so downstream pipelines will not break between runs.
How often is the data refreshed?
Every run reads the live source. There is no cache, so the data is as current as the website itself.
What if the site changes its layout?
Open an issue on the Issues tab and it gets fixed โ the Actor is actively maintained.
Is it legal?
This Actor reads only publicly available pages on IRS Tax โ the same content any visitor sees without logging in. It does not bypass authentication and does not collect private data. You remain responsible for how you use the output: respect the source's terms of service, applicable copyright, and data-protection law such as GDPR where personal data is involved.
Related scrapers
Browse the full collection at apify.com/logiover โ Lead Generation, Other tools and more.