Scholarship Finder Scraper — 23,000+ US Scholarships, 23 Fields avatar

Scholarship Finder Scraper — 23,000+ US Scholarships, 23 Fields

Pricing

from $3.00 / 1,000 results

Go to Apify Store
Scholarship Finder Scraper — 23,000+ US Scholarships, 23 Fields

Scholarship Finder Scraper — 23,000+ US Scholarships, 23 Fields

Find 23,000+ US scholarships — awards, deadlines, eligibility, amounts. 23 fields. Keyword search, sort by award/deadline. Education financial aid research.

Pricing

from $3.00 / 1,000 results

Rating

0.0

(0)

Developer

qingwa

qingwa

Maintained by Community

Actor stats

1

Bookmarked

15

Total users

1

Monthly active users

9 days ago

Last modified

Share

Scholarship Finder Scraper

Extract comprehensive US scholarship data from CollegeScholarships.org — 23,000+ scholarships with 20+ fields per record.

Features

  • 🎓 23,000+ scholarships from CollegeScholarships.org
  • 📊 20+ fields per scholarship (award min/max, deadline, eligibility, sponsor, etc.)
  • 🔍 Keyword search — server-side filter by title, description, or major
  • 💰 Award sorting — sort by highest/lowest award or deadline (chronological)
  • 🏷️ Min award filter — applied before detail fetching, no wasted requests
  • 📄 Detail page enrichment — concurrent fetching for full data
  • 🚀 Fast & reliable — pure HTTP, no JavaScript rendering needed

Output Fields

FieldDescriptionSource
titleScholarship nameListing
detail_urlFull URL to scholarship pageListing
awardAward amount (text)Listing
award_amountAward amount (numeric)Listing
deadlineApplication deadlineListing
descriptionScholarship descriptionListing
geographic_restrictionsGeographic eligibilityListing
enrollment_levelRequired enrollment levelListing
majorsEligible majors/fieldsListing
min_awardMinimum award amountDetail
max_awardMaximum award amountDetail
award_typeType (Scholarship/Grant/Fellowship)Detail
renewableWhether award is renewableDetail
awarded_annuallyWhether awarded every yearDetail
repay_requiredWhether repayment is requiredDetail
unlimited_awardsWhether awards are unlimitedDetail
enrollment_detailDetailed enrollment requirementsDetail
ethnicityEthnicity requirementsDetail
genderGender requirementsDetail
requirementsAdditional requirementsDetail
gpaGPA requirementDetail
sponsor_nameScholarship sponsorDetail
sponsor_urlSponsor websiteDetail
sponsor_phoneSponsor phoneDetail
sponsor_addressSponsor addressDetail
apply_urlDirect application linkDetail
full_descriptionFull description from detail pageDetail
sourceData source identifier—

Input Parameters

ParameterTypeDefaultDescription
keywordString—Filter by keyword (e.g., "engineering", "nursing"). Uses the site's own search when provided
maxResultsSelect100Max results: 50/100/200/500/1000/2000/5000
maxPagesSelect10Max pages to scrape (30/page): 5/10/20/50/100/769
fetchDetailsCheckbox✓Fetch detail pages for enriched data
minAwardSelect—Only scholarships with award ≥ this value (applied before detail fetching)
requestConcurrencySelect8Parallel detail-page fetches: 3/5/8/12
sortBySelectnoneSort by: none / award_high / award_low / deadline (chronological)

Changelog

v1.1 (2026-09-28)

  • Concurrent detail fetching: detail pages now fetched in parallel (configurable via requestConcurrency, default 8) instead of serially with a 1s sleep — 5 detail pages complete in ~2s instead of ~10s+
  • Server-side keyword/award search: when keyword or minAward is set, the actor POSTs the site's own search form (with CSRF token) so listing pages arrive pre-filtered — no more fetching thousands of pages just to filter locally
  • Concurrent listing pages: up to 4 listing pages fetched in parallel instead of strictly serial with 1.5s sleeps
  • Correct deadline sorting: sortBy=deadline now parses real month/day (February before April); vague deadlines ("Rolling", "Varies") sort last
  • Dedup by detail_url: previously deduped by title only, which dropped distinct scholarships sharing a name
  • minAward applied before detail fetching: previously filtered after all detail pages were fetched
  • Retries with backoff for transient HTTP statuses (429/500/502/503/504)
  • Hardened inputs: invalid maxResults/maxPages/requestConcurrency values fall back to defaults instead of crashing
  • Fixed httpx crash on bracketed IPv6 entries in NO_PROXY (sandbox environments)
  • Added output schema + dataset schema; removed committed __pycache__ artifacts from the source bundle
  • New optional input requestConcurrency (3/5/8/12)

v1.0 (2026-09-28)

  • Initial release: listing + detail scraping from CollegeScholarships.org, keyword filter, award/deadline sorting, min-award filter.