Openverse Open-License Media Scraper
Pricing
from $13.00 / 1,000 result items
Openverse Open-License Media Scraper
Search 800M+ openly licensed images, audio clips and graphics across Flickr, Wikimedia, Europeana, Smithsonian, NASA and 50+ CC and public-domain providers. Returns title, creator, license, attribution, source URL, file size, dimensions, tags and direct media URL. Filter by license or source.
Pricing
from $13.00 / 1,000 result items
Rating
0.0
(0)
Developer
ParseForge
Maintained by CommunityActor stats
0
Bookmarked
18
Total users
3
Monthly active users
4 days ago
Last modified
Share

๐จ Openverse Media Scraper
๐ Search 800M+ openly licensed images, audio, and graphics across 50+ providers.
The Openverse Media Scraper searches WordPress.org's Openverse index of openly licensed media and returns structured records for images, audio clips, illustrations, and graphics. Every result is licensed under Creative Commons or in the public domain, with full attribution metadata.
The catalog aggregates 800M+ items across 50+ providers (Flickr, Wikimedia Commons, Europeana, Smithsonian, NASA, Bio Diversity Library, Rawpixel). Filters run server-side, so a single run can isolate CC0 sunsets, Smithsonian sketches, or NASA imagery only.
| ๐ฏ Target Audience | ๐ก Primary Use Cases |
|---|---|
| Content creators, designers, educators, marketing teams, journalists, app developers, AI training pipelines | Content libraries, blog illustrations, social media assets, AI training datasets, educational materials |
๐ What the Openverse Media Scraper does
Five filtering workflows in a single run:
- ๐ Keyword search. Match titles, descriptions, tags, and creator names across the catalog.
- ๐ท๏ธ License filter. Restrict by CC license (CC0, CC-BY, CC-BY-SA) or public domain.
- ๐ Source filter. Restrict to one provider.
- ๐ Aspect ratio. Tall, wide, or square (images only).
- ๐ต Media type toggle. Switch between images and audio.
๐ก Why it matters: clean, server-side filtering removes the parser-and-pagination work from your team and keeps your dataset fresh on every run.
๐ Data fields
Each record includes: creator, fileType, height, id, license, licenseVersion, openverseUrl, source, sourceUrl, thumbnailUrl, title, width. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.
๐ How to use
- ๐ Sign up. Create a free account with $5 credit (takes 2 minutes).
- ๐ Open the Actor. Go to the Openverse Media Scraper page on the Apify Store.
- ๐ฏ Set input. Pick your filters and
maxItems. - ๐ Run it. Click Start and let the Actor collect your data.
- ๐ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.
โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.
๐ Recommended Actors
- ๐ Project Gutenberg Books - 75,000+ free public-domain books
- ๐ Open Library Books - 30M+ books and editions
- ๐จ Met Museum Scraper - Metropolitan Museum public-domain artworks
- ๐ Wikidata Entity Search - 100M+ open knowledge-graph entities
- ๐ฌ TVMaze TV Shows - TV show metadata and episodes
๐ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.
โ ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by WordPress.org, Openverse, or any of the upstream content providers. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.
๐ Need Help?
If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.
For faster answers, join our Discord. It's the best place to get support and suggest new actors.