Openverse Open-License Media Scraper avatar

Openverse Open-License Media Scraper

Pricing

from $13.00 / 1,000 result items

Go to Apify Store
Openverse Open-License Media Scraper

Openverse Open-License Media Scraper

Search 800M+ openly licensed images, audio clips and graphics across Flickr, Wikimedia, Europeana, Smithsonian, NASA and 50+ CC and public-domain providers. Returns title, creator, license, attribution, source URL, file size, dimensions, tags and direct media URL. Filter by license or source.

Pricing

from $13.00 / 1,000 result items

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

0

Bookmarked

18

Total users

3

Monthly active users

4 days ago

Last modified

Share

ParseForge Banner

๐ŸŽจ Openverse Media Scraper

๐Ÿš€ Search 800M+ openly licensed images, audio, and graphics across 50+ providers.

The Openverse Media Scraper searches WordPress.org's Openverse index of openly licensed media and returns structured records for images, audio clips, illustrations, and graphics. Every result is licensed under Creative Commons or in the public domain, with full attribution metadata.

The catalog aggregates 800M+ items across 50+ providers (Flickr, Wikimedia Commons, Europeana, Smithsonian, NASA, Bio Diversity Library, Rawpixel). Filters run server-side, so a single run can isolate CC0 sunsets, Smithsonian sketches, or NASA imagery only.

๐ŸŽฏ Target Audience๐Ÿ’ก Primary Use Cases
Content creators, designers, educators, marketing teams, journalists, app developers, AI training pipelinesContent libraries, blog illustrations, social media assets, AI training datasets, educational materials

๐Ÿ“‹ What the Openverse Media Scraper does

Five filtering workflows in a single run:

  • ๐Ÿ” Keyword search. Match titles, descriptions, tags, and creator names across the catalog.
  • ๐Ÿท๏ธ License filter. Restrict by CC license (CC0, CC-BY, CC-BY-SA) or public domain.
  • ๐Ÿ“ Source filter. Restrict to one provider.
  • ๐Ÿ“ Aspect ratio. Tall, wide, or square (images only).
  • ๐ŸŽต Media type toggle. Switch between images and audio.

๐Ÿ’ก Why it matters: clean, server-side filtering removes the parser-and-pagination work from your team and keeps your dataset fresh on every run.

๐Ÿ“Š Data fields

Each record includes: creator, fileType, height, id, license, licenseVersion, openverseUrl, source, sourceUrl, thumbnailUrl, title, width. These field names come straight from the actor's dataset schema, so what you see here is what lands in your dataset.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Go to the Openverse Media Scraper page on the Apify Store.
  3. ๐ŸŽฏ Set input. Pick your filters and maxItems.
  4. ๐Ÿš€ Run it. Click Start and let the Actor collect your data.
  5. ๐Ÿ“ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.

โš ๏ธ Disclaimer: this Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by WordPress.org, Openverse, or any of the upstream content providers. All trademarks mentioned are the property of their respective owners. Only publicly available open data is collected.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.