Go to example tasks
Check which sitemaps a website actually publishes
One row per sitemap file: where it was found - robots.txt, a common path, or a parent index - whether it answered, how many URLs it contributed, and a note when something was filtered, duplicated or cut short. This is the view to open when a crawl comes back short.
Sitemap URL Extractor — All Page URLs from a Websitepower_on/sitemap-url-list
Website
Record type
Sitemap
Found via
+4 fieldsTextNumberBooleanListObject
Input
Websites or sitemap URLs(required):https://www.allbirds.com+2
Report only - count the pages, do not return them:true
Max sitemap files per website:200
Output fields
Website
Record type
Sitemap
Found via
Status
URLs
Child sitemaps
Note
Sign up on Apify01
Create your Apify account to access the Sitemap URL Extractor — All Page URLs from a Website.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
