Internet Archive Media & Video Finder avatar

Internet Archive Media & Video Finder

Pricing

from $1.00 / 1,000 dataset items

Go to Apify Store
Internet Archive Media & Video Finder

Internet Archive Media & Video Finder

Keyless Internet Archive media search: 17M+ videos, audio & text collections by topic or collection (e.g. Prelinger public-domain reels). Resolved download URL + format + runtime per item, plus a license field and public-domain filter so content is safe to reuse. More than web-snapshot scrapers.

Pricing

from $1.00 / 1,000 dataset items

Rating

0.0

(0)

Developer

jedi solana

jedi solana

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Categories

Share

Keyless Internet Archive media search: 17M+ videos, audio & text collections by topic or collection (e.g. Prelinger public-domain reels). Resolved download URL + format + runtime per item, plus a license field and public-domain filter so content is safe to reuse. More than web-snapshot scrapers.

What you get

Every run writes a dataset with one row per record:

{
"title": "Wartime Nutrition",
"identifier": "WartimeN1943",
"creator": "U.S. Office of War Information",
"description": "Wartime work of public welfare agencies in the field of nutrition.",
"downloads": 962298,
"url": "https://archive.org/details/WartimeN1943",
"mediatype": "movies",
"license": "http://creativecommons.org/licenses/publicdomain/",
"open_license": true,
"date": "1943-01-01T00:00:00Z"
}

Input

{
"query": "public domain",
"collection": "prelinger",
"mediatype": "movies",
"minDownloads": 1,
"publicDomainOnly": false,
"since": "",
"maxItems": 10
}
FieldTypeDescription
querystringQuery: the query parameter. Default: "public domain".
collectionstringCollection: the collection parameter. Default: "prelinger".
mediatypestringMediatype: the mediatype parameter. Default: "movies".
minDownloadsintegerMinDownloads: the mindownloads parameter. Default: 1.
publicDomainOnlybooleanPublicDomainOnly: the publicdomainonly parameter. Default: false.
sincestringSince: the since parameter. Default: "".
maxItemsintegerMaxItems: the maxitems parameter. Default: 10.

Use cases

  • Monitor on a schedule — run this Actor on a timer and get a fresh, normalized snapshot every time; diff the datasets to see what changed.
  • Feed a dashboard or model — clean, one-JSON-item-per-record output is ready to pipe into a spreadsheet, database, or LLM prompt without any post-processing.
  • Research & due diligence — pull a structured set of records for a topic, name, or window without scraping the public source by hand.

Limits & notes

  • No API key, no login: it only reads public endpoints/feeds. If the upstream source is degraded, the run still succeeds and returns an honest status item instead of failing.
  • Results reflect what the public source returned at run time (top-N windows, newest-first where applicable).
  • Pay-per-event via synthetic events: each default-dataset item is billed through apify-default-dataset-item ($0.001 per item), plus apify-actor-start.
  • Free to try on the free plan; PPE pricing applies to paid-plan users.

Categories

VIDEOS, OTHER