SBIR & STTR Award Scraper - Funded R&D Company Leads avatar

SBIR & STTR Award Scraper - Funded R&D Company Leads

Pricing

from $3.31 / 1,000 sbir/sttr award / company leads

Go to Apify Store
SBIR & STTR Award Scraper - Funded R&D Company Leads

SBIR & STTR Award Scraper - Funded R&D Company Leads

Scrape the official SBIR.gov award database (SBIR + STTR): company, UEI, website, address, business contact + principal investigator emails & phones, award amount, agency, phase & set-asides. Current 2025-2026 awards. One row per award or per company. B2B leads + monitoring.

Pricing

from $3.31 / 1,000 sbir/sttr award / company leads

Rating

0.0

(0)

Developer

Scrape Sage

Scrape Sage

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

12 days ago

Last modified

Share

Disclaimer: This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the U.S. Small Business Administration (SBA) or any government body. All trademarks mentioned are the property of their respective owners. "SBIR.gov" is referenced only to describe the public data source this Actor collects from.

Extract the complete SBIR & STTR award database from the official SBIR.gov open data — every federal Small Business Innovation Research (SBIR) and Small Business Technology Transfer (STTR) award, with the funded company, the award economics, and the decision-maker contacts that make each award a ready-to-use B2B lead.

Every record carries the company name, website and full address, the award (program, phase, agency, branch, dollar amount, topic, dates), and two named contacts — the business/company contact and the Principal Investigator — each with title, email and phone, plus the partner research institution (for STTR), set-aside ownership flags, the full project abstract and a derived lead score.

No login, no cookies, no browser, no API key - fast, reliable extraction straight from the government bulk dataset.

Current data, not an archive. This actor reads the SBIR.gov bulk file that SBIR.gov actually keeps refreshing (monthly), so a default run returns 2025 and 2026 awards - and every record carries the company's UEI, the federal entity identifier that replaced DUNS in 2022. Every run logs the source file's exact publication date and repeats it in each record's dataAsOf field, so you always know how fresh your leads are.

Two ways to shape the output. Ask for one row per award (default), or set recordType: "companies" to get one row per funded company with its whole award history rolled up - total awards, total funding, every agency and phase, first and latest award year, and the best contact found across all of them. Both are billed at the same single price.

Why this SBIR/STTR scraper?

A company that just won federal R&D money is a high-intent B2B lead: it's funded, it's hiring, and it's buying. Generic exporters dump a few API fields and miss the contact data that turns an award into outreach. This actor ships the richest record in the category — the full firmographics and the named business contact and the Principal Investigator, with emails and phones for both.

DataGeneric SBIR exportersThis actor
Company, agency, phase, amount, year
Company website + full mailing addresspartial
Business contact name, title, email, phone
Principal Investigator name, title, email, phonepartial
Partner research institution + POC (STTR)
Set-aside flags (woman-owned / HUBZone / disadvantaged)partial
Employees, DUNS, topic code, solicitation, contract #partial
Full project abstract
Lead score (contactability + recency + phase + size)
Filter by agency, phase, state, year, amount, keywordpartial
Monitoring mode — only newly-funded companies

Use cases

  • B2B sales & lead generation — newly-funded R&D companies are in-market buyers. Export them with email, phone, website and award size and drop them straight into your CRM or outreach sequence.
  • Government-contracting business development — find subcontractors, teaming partners and primes by agency, branch, topic and technology.
  • Recruiting & talent — Phase II winners are scaling and hiring engineers and scientists; the PI is the technical decision-maker.
  • Investors & M&A — track non-dilutive-funded deep-tech startups by technology, geography and award trajectory.
  • Professional services — IP/patent attorneys, accountants, grant consultants and R&D tax-credit firms can target firms by funding stage and agency.
  • Market & competitive research — map who is winning R&D funding in AI, batteries, biotech, defense, space and more, by agency and over time.

How to use

  1. Sign up for Apify — the free plan is enough to try this actor.
  2. Open the SBIR & STTR Award Scraper, set your filters (e.g. agency Defense, state MA, awardYearFrom: 2023, withEmail: true), and pick how many results you want.
  3. Click Start and watch results stream into the dataset table.
  4. Export as JSON, CSV, Excel, XML or RSS — or pull results programmatically via the Apify API.

Input

{
"agencies": ["Defense"],
"states": ["MA"],
"phases": ["Phase II"],
"awardYearFrom": 2023,
"withEmail": true,
"sortBy": "newest",
"maxResults": 500
}

All filters are optional and combine with AND. Highlights:

  • recordTypeawards (default, one row per award) or companies (one row per funded company, every award rolled up).
  • programsSBIR, STTR, or both (empty = both).
  • phasesPhase I (feasibility) and/or Phase II (development; proven, larger awards).
  • agencies — match the funding agency by keyword: Defense, Health (HHS/NIH), Energy, Homeland Security, Aeronautics (NASA), Science Foundation (NSF), Agriculture, Commerce, Education, Transportation, Environmental.
  • branches — sub-agency contains: Navy, Air Force, Army, DARPA, NIH, ARPA-E, NOAA, NIST
  • keywords — match any term in the title or abstract (e.g. artificial intelligence, battery, hypersonic, quantum) to build a technology-specific list.
  • states / cities / zipCodes — locate companies geographically. states accepts either form: "MA" or "Massachusetts".
  • awardYearFrom / awardYearTo, minAwardAmount / maxAwardAmount, awardedAfter / awardedBefore — funding window and dollar size.
  • minEmployees / maxEmployees, womenOwnedOnly / hubZoneOwnedOnly / disadvantagedOnly, researchInstitutionQuery — company size, set-aside status and STTR partner.
  • withEmail / withPhone / withWebsite — keep only contactable records.
  • sortBynewest (default), awardAmountHigh, awardAmountLow, oldest. Every mode scans the whole file and sorts globally, so newest really is the newest.
  • includeAbstract — off by default; see Data freshness below before turning it on.
  • monitorMode / monitorKey — return only awards new since the last run (see below).
  • maxResults, deduplicateResults, proxyConfiguration.

Tip: run with no filters at all and you get a small 25-record sample rather than a full-price bulk pull, so an exploratory or agent-issued call is never an expensive surprise. Set any filter - or your own maxResults - for a full run.

Data freshness - which SBIR.gov file this reads

SBIR.gov publishes two bulk award files, and they are not equally current:

FileSizeRefreshedNewest awardsUEIAbstract
Current (this actor's default)~87 MBmonthly2026yesno
Archival (includeAbstract: true)~351 MBfrozen since 2026-01-012024noyes

Only the archival file carries the full project abstract, and SBIR.gov has not refreshed it since 1 January 2026 - its newest awards are from 2024. So:

  • Leave includeAbstract off (the default) for current, contactable leads. You get 2025/2026 awards and UEI.
  • Turn it on only when you specifically need abstract text - for keyword research across award descriptions, say - and can accept awards up to ~2024. The run log warns you, and the run's status message records which file answered.
  • Every record's dataAsOf field is the source file's own Last-Modified date, and the run warns on the card whenever SBIR.gov's file is more than ~10 weeks old. That is a source-side publication lag, never a scraping failure.

keywords searches the award title always, and the abstract as well when includeAbstract is on.

Output

By default you get one clean, dense table — every column applies to every row. The dataset ships four ready-made views: Awards, Company leads, Contacts (business + PI) and Funded companies (for recordType: "companies").

Award record - 61 fields. Company record - 54 fields: company, uei, duns, totalAwards, totalAwardAmount, totalAwardTier, agencies[], branches[], programs[], phases[], topicCodes[], firstAwardYear, latestAwardYear, the latest award snapshot, full address, ownership flags, the best contact across the whole portfolio, and a company leadScore.

An award record:

{
"recordType": "award",
"company": "ATA ENGINEERING, INC.",
"companyWebsite": "http://www.ata-e.com",
"duns": "...",
"numberEmployees": 205,
"womenOwned": false,
"hubZoneOwned": false,
"sociallyEconomicallyDisadvantaged": false,
"setAsides": [],
"address1": "...",
"city": "San Diego",
"state": "CA",
"zip": "92128-4695",
"zipCode": "92128",
"country": "US",
"fullAddress": "..., San Diego, CA, 92128-4695",
"program": "STTR",
"phase": "Phase II",
"phaseNumber": 2,
"agency": "Department of Defense",
"branch": "Navy",
"awardTitle": "…",
"awardYear": 2024,
"awardAmount": 997409,
"awardAmountText": "997,409",
"awardTier": "mid",
"contractNumber": "N68335-24-C-0081",
"agencyTrackingNumber": "…",
"topicCode": "N22A-T016",
"solicitationNumber": "…",
"solicitationYear": 2022,
"proposalAwardDate": "2023-10-23",
"contractEndDate": "2025-10-30",
"dateOfNotification": "2023-05-05",
"contactName": "Heather Wilkens",
"contactTitle": "…",
"contactPhone": "(858) 480-2043",
"contactEmail": "hwilkens@ata-e.com",
"piName": "Timothy Palmer",
"piTitle": "…",
"piPhone": "(585) 480-2066",
"piEmail": "tim.palmer@ata-e.com",
"primaryEmail": "hwilkens@ata-e.com",
"primaryPhone": "(858) 480-2043",
"researchInstitution": "University of Arkansas",
"riPocName": "…",
"riPocPhone": "…",
"abstract": "…full project abstract…",
"hasEmail": true,
"hasPhone": true,
"hasWebsite": true,
"isPhaseII": true,
"isSttr": true,
"uei": "SATYSBWG3FL7",
"stateName": "Massachusetts",
"leadScore": 93,
"dataAsOf": "2026-08-01",
"scrapedAt": "2026-06-20T13:00:00.000Z"
}

What to expect (field coverage)

SBIR.gov is government-entered data. Award economics, company and location are essentially always present; contact richness is highest for recent awards (which is exactly what sortBy: "newest" returns). Older awards (1980s–2000s) often have placeholder or missing contacts — those are cleaned to null, never faked.

Measured on a live run (Department of Defense, Massachusetts, 25 most recent awards, 2026-08-31):

Field groupCoverage
Company, agency, program, phase, amount, year, address100%
UEI100%
Employees100%
Business contact email / Principal Investigator email / phone100%
Company website96%
DUNS96%
Research institutionpresent on STTR awards only
Abstractonly when includeAbstract is on (archival file)

Coverage is highest on recent awards, which is what a default run returns. Older awards (1980s-2000s) often carry placeholder or missing contacts - those are cleaned to null, never faked.

A blank field means SBIR.gov didn't publish it for that award — not that scraping failed. Nothing is dropped, so you always get the richest dataset available.

Monitoring mode — only newly-funded companies

Turn on monitorMode and the actor remembers every award it has already returned (in a named key-value store) and, on the next run, emits only awards that are new since last time — each tagged monitorEvent: "new".

This is orthogonal to Apify Schedules: the Schedule starts the run on your cadence (say, weekly); monitoring mode decides what's new. Use a distinct monitorKey per saved watch (e.g. one per agency or state) so different monitors keep separate memory. The result: a clean, de-duplicated feed of newly-funded R&D companies dropped straight into your CRM. Because SBIR.gov refreshes the source file monthly, a monthly or weekly schedule matches the data's own cadence.

How much does it cost to scrape SBIR.gov?

This Actor uses Apify's pay-per-event pricing: you are charged only for the results it delivers, with no monthly rental and no start fee. The events it can charge are:

  • SBIR/STTR award / company lead - One SBIR or STTR federal R&D award record from the official SBIR.gov database: the funded company (name, website, full address, employees, set-aside flags), the award (program, phase, agency, branch, dollar amount, topic code, dates), the business contact AND principal investigator (names, titles, emails, phones), the partner research institution, the full abstract and a derived lead score.

The current price of each event is shown on the Pricing tab of this page. Set a maximum total charge on the run if you want a hard cap on spend, and use the input limits to control how much the Actor fetches.

Automate & schedule

Run this actor on autopilot and pull results into your own stack:

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: 'MY_APIFY_TOKEN' });
const run = await client.actor('scrapesage/sbir-sttr-award-scraper').call({
agencies: ['Defense'],
states: ['MA'],
awardYearFrom: 2023,
withEmail: true,
maxResults: 500,
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(`Got ${items.length} SBIR/STTR awards`);

Integrate with any app

Connect the dataset to 5,000+ apps — no code required:

  • Make — multi-step automation scenarios.
  • Zapier — push new award leads straight into your CRM.
  • Slack — get notified when a monitored agency or state has new awards.
  • Google Drive / Sheets — auto-export every run to a spreadsheet.
  • Airbyte — pipe results into your data warehouse.
  • GitHub — trigger runs from commits or releases.

Use with AI assistants (MCP)

The output is clean, LLM-ready JSON. Call this actor from Claude, ChatGPT, or any agent framework through the Apify MCP server — ask your assistant to "list every Phase II SBIR award in Massachusetts since 2023 with the company email" and let it run the scraper for you.

Agent-ready: autonomous payments (x402 & Skyfire)

This actor is agent-ready — AI agents can discover it, run it, and pay for it autonomously, with no Apify account and no human in the loop. It uses pay-per-event pricing and limited permissions, so it qualifies for Apify's agentic-payment standards:

  • x402 — an open, HTTP-native payment protocol. Agents pay per run in USDC on the Base network directly through the Apify MCP server — no account, no API key.
  • Skyfire — agent-to-service payments for fully autonomous AI-agent workflows.

Building an AI agent, MCP tool, or autonomous data pipeline? This scraper is ready to plug in and pay as it goes.

More US-B2B & funding-data scrapers from scrapesage

Build a complete funded-company & federal-money lead-gen stack:

FAQ

Where does the data come from? The official SBIR.gov public award dataset (every SBIR & STTR award, all 11 participating agencies, going back to the 1980s), downloaded directly as the government bulk file.

Does it need the SBIR API or a key? No. This actor reads the public bulk award file directly — no key, no login, no browser — which is also far more reliable than the rate-limited public API.

Why are some old awards missing emails/phones? Contact fields are richest on recent awards. SBIR.gov stores placeholder values (e.g. () -) for missing contacts on older records; the actor cleans those to null rather than emitting junk. Use sortBy: "newest" and withEmail: true for the cleanest lead list.

What's the difference between SBIR and STTR? Both fund small-business R&D. STTR additionally requires a partnership with a research institution (university or federal lab), which appears in the researchInstitution field.

Can I export to Google Sheets, CSV or Excel? Yes — one click in the dataset view, or automatically on every run via the Google Drive integration.

How do I get only newly-funded companies? Turn on monitorMode and run on a Schedule — you'll get only awards added since the last run.

Is scraping this legal? This actor collects publicly available U.S. government open data. You're responsible for using the data in compliance with applicable laws (e.g. CAN-SPAM/CCPA for outreach).

Data & lawful use

The source is an official public register, published so that anyone can consult it. Records can name individuals (principal investigators and company contacts), so the output may contain personal data even though it is public. If you are in the EU or UK you are the data controller for what you do with it: have a lawful basis, honour access and deletion requests, and respect the register's own reuse conditions, which can restrict marketing or commercial use.

Under Apify's Standard Actor Contract, which governs your use of this Actor, you are the controller of any personal data in your input and output and scrapesage acts only as your processor: that data is processed solely to run your job, written only to your own Apify storage, never used for any other purpose and never shared onward. If you need help with a data-subject request that involves this Actor's output, open an issue on the Issues tab.

Disclaimer

This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by the U.S. Small Business Administration (SBA) or any government body. All trademarks mentioned are the property of their respective owners.

"SBIR.gov" is referenced only in a descriptive, nominative sense - to identify the public data source this Actor collects from. This Actor is not an official product or service of the U.S. Small Business Administration (SBA) and is not authorised or certified by it. It collects only publicly available records; you are responsible for ensuring your use of that data complies with applicable laws, regulations and the source's own terms of use or reuse conditions.

Need help?

Open an issue on the actor's Issues tab, or visit the Apify help center. Feature requests are welcome — this actor is actively maintained.