Puma Product Scraper
Pricing
from $0.99 / 1,000 results
Puma Product Scraper
Extract Puma product data at scale. Scrape prices, descriptions, images, and reviews from Puma.com. Real-time monitoring, zero blocks. Perfect for price tracking, competitive analysis, and inventory management.
Pricing
from $0.99 / 1,000 results
Rating
5.0
(3)
Developer
Shahid Irfan
Maintained by CommunityActor stats
1
Bookmarked
12
Total users
0
Monthly active users
10 days ago
Last modified
Categories
Share
What does Puma Product Scraper do?
Puma Product Scraper is a Puma product data extractor for collecting product listings from public Puma category, search, tag, and country-level listing URLs. Enter one Puma listing URL, choose how many products and pagination pages to process, and receive structured product records with identifiers, names, brands, colors, prices, ratings, reviews, availability signals, images, sizes, promotions, and product links.
The Actor is useful for ecommerce market research, competitor price tracking, catalog monitoring, merchandising analysis, and repeatable product snapshots. Results are saved to an Apify dataset and can be reviewed, exported, scheduled, or connected to other workflows.
Why use Puma Product Scraper?
- Product catalog research - Build a structured view of Puma products from a category or search listing.
- Price monitoring - Compare regular, sale, promotion, and best-price values across repeated runs.
- Merchandising analysis - Review colors, images, badges, promotions, sizes, and orderability signals.
- Review and product intelligence - Track average ratings and review counts when Puma publishes them.
- Country-aware collection - Use Puma URLs from supported country and language storefronts; the URL determines the listing context.
- Automation-ready datasets - Download JSON, CSV, Excel, XML, and other formats supported by Apify, or connect the dataset to integrations and scheduled workflows.
What data can you extract from Puma?
Each dataset item represents a product variant or listing record returned from the requested Puma listing context. The Actor removes null, blank, and empty nested values, so optional fields may not appear in every item.
| Field | Type | Description |
|---|---|---|
source | String | Source label, normally puma. |
source_mode | String | Collection mode, such as category or search. |
category_url | String | Puma URL supplied in start_url. |
category_path | String | Active listing path used for the record. |
category_id | String | Puma category identifier when available. |
category_name | String | Puma category name when available. |
search_term | String | Search term detected from a search URL, when applicable. |
page_offset | Number | Listing offset from which the record was collected. |
product_id | String | Product or variant identifier from the listing. |
variant_id | String | Puma variant identifier. |
master_id | String | Parent or master product identifier. |
sku | String | Variant SKU or product identifier used by the listing. |
ean | String | EAN identifier when available. |
product_url | String | Direct Puma product page URL. |
name | String | Product name. |
header | String | Product header text when available. |
sub_header | String | Product sub-header text when available. |
brand | String | Brand label returned for the product. |
color_name | String | Human-readable variant color name. |
color_code | String | Puma color value or code. |
orderable | Boolean | Whether the returned variant is marked orderable. |
orderable_color_count | Number | Number of orderable colors when available. |
app_exclusive | Boolean | App-exclusive indicator when supplied by Puma. |
app_only_from | String | App-only start timestamp when available. |
app_only_to | String | App-only end timestamp when available. |
price_regular | Number | Regular product price. |
price_sale | Number | Sale price when available. |
price_promotion | Number | Promotion price when available. |
price_best | Number | Best price value when available. |
price_tax | Number | Tax amount when available. |
price_tax_rate | Number | Tax rate when available. |
is_sale_price_elapsed | Boolean | Whether Puma marks the sale price as elapsed. |
rating | Number | Average product rating. |
review_count | Number | Number of product reviews. |
score_rating | Number | Additional score rating when available. |
score_amount | Number | Additional score amount when available. |
badge_labels | Array | Product badge labels. |
variant_promotion_messages | Array | Promotion messages attached to the variant. |
master_promotion_messages | Array | Promotion messages attached to the parent product. |
image_url | String | Primary product image URL. |
image_vertical_url | String | Vertical version of the primary image when available. |
image_alt | String | Alternative text for the primary image. |
all_image_urls | Array | Available image URLs for the variant. |
sizes | Array | Size groups with labels, availability, product IDs, and maximum orderable quantities. |
color_options | Array | Other color options with names, values, and image URLs. |
scraped_at | String | ISO timestamp for when the record was collected. |
How to scrape Puma product data
- Open Puma Product Scraper on Apify.
- Enter a public Puma category, search, tag, or country-level listing URL in
start_url. - Set
results_wantedandmax_pagesfor the size of the collection. - Optionally configure
proxyConfigurationif the source connection is rate-limited. - Start the run and review the dataset preview.
- Export the dataset or connect it to your ecommerce research and monitoring workflow.
The Actor accepts one starting URL per run. It follows listing pagination up to the configured page limit and stops when the requested result count is reached, no more products are returned, or the page limit is reached.
Input Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
start_url | String | No | https://us.puma.com/us/en/men/shoes | Puma category, search, tag, or country-level listing URL. The URL also determines the country and language context when those are included in the URL. |
results_wanted | Integer | No | 20 | Maximum number of product records to save. Must be at least 1. |
max_pages | Integer | No | 1 | Maximum number of listing pagination pages to process. Must be at least 1. |
proxyConfiguration | Object | No | {"useApifyProxy": false} | Optional Apify Proxy configuration for connections that are rate-limited. |
The actor input schema does not provide a multiple-URL array, keyword field, price filter, brand filter, or configurable page-size field. Use the appropriate Puma listing URL to express the category or search target.
Usage Examples
Basic Category Extraction
Collect up to 20 products from Puma's men's shoes category:
{"start_url": "https://us.puma.com/us/en/men/shoes","results_wanted": 20}
Search URL Collection
Collect products returned by a Puma search URL. Search parameters remain part of the URL:
{"start_url": "https://us.puma.com/us/en/search?q=running","results_wanted": 50,"max_pages": 3}
Larger Run with Apify Proxy
Process more listing pages and enable the optional proxy configuration when repeated requests need additional connection support:
{"start_url": "https://de.puma.com/de/de/herren/schuhe","results_wanted": 100,"max_pages": 5,"proxyConfiguration": {"useApifyProxy": true}}
Sample Output
This illustrative dataset item shows the main fields returned for a Puma product variant. Optional fields are omitted when Puma does not provide them:
{"source": "puma","source_mode": "category","category_url": "https://us.puma.com/us/en/men/shoes","category_path": "/men/shoes","category_id": "mens-shoes","category_name": "Men's Shoes and Sneakers","page_offset": 0,"product_id": "198556306179","variant_id": "198556306179","master_id": "308762","sku": "198556306179","product_url": "https://us.puma.com/us/en/pd/scuderia-ferrari-trinity-2-mens-sneakers/308762?swatch=07","name": "Scuderia Ferrari Trinity 2 Men's Sneakers","brand": "Ferrari","color_name": "PUMA Black-Speed Yellow","color_code": "07","orderable": true,"price_regular": 103,"price_sale": 103,"rating": 0,"review_count": 0,"badge_labels": ["New"],"image_url": "https://images.puma.com/image/upload/f_auto,q_auto,b_rgb:fafafa,w_2000,h_2000/global/308762/07/sv01/fnd/PNA/fmt/png/Scuderia-Ferrari-Trinity-2-Men's-Sneakers","scraped_at": "2026-04-01T10:55:22.102Z"}
Data-quality behavior and limitations
- Empty values are omitted - Null, blank, and empty nested values are removed from each dataset item. A missing key generally means the source did not provide a value for that product.
- Pricing and review recovery - When listing data contains missing or zero price or review values, the Actor attempts to recover them from available product-level data. Recovery is best effort and the original value may remain missing or zero.
- Variant-level records - The dataset is based on listing variants. Similar products in different colors or variants can therefore appear as separate records when Puma returns them separately.
- Duplicate control - Records are deduplicated using variant or product identifiers during a run.
- Result counts are maximums -
results_wantedis a cap, not a guarantee. The final dataset can contain fewer records if the listing has fewer products, pagination ends, records cannot be mapped, ormax_pagesis reached. - Source-dependent fields - Ratings, reviews, prices, promotions, sizes, images, and orderability can vary by country, language, product, and storefront state.
- Listing-focused collection - The supported workflow starts from Puma listing URLs. The input does not accept a list of product detail URLs or provide separate detail-page filters.
- Public-page changes - Puma can change listing availability, URLs, fields, or storefront behavior. Check a small run after changing URL patterns and report persistent issues through the Actor page.
Tips for Best Results
- Use a complete Puma listing URL that already represents the category or search you want to analyze.
- Start with
results_wanted: 20andmax_pages: 1to validate the dataset before increasing the run size. - Use
max_pagesto control how far pagination proceeds; the Actor uses a fixed internal page request size. - For country-specific research, use the relevant Puma country and language URL rather than relying on a different storefront.
- Treat absent optional fields as source-data limitations instead of assuming the entire record failed.
- For recurring price or catalog monitoring, compare datasets from scheduled runs using stable identifiers such as
product_id,variant_id,master_id, andsku.
Integrations and Export Formats
- Google Sheets - Review product prices, ratings, and assortment changes in a spreadsheet.
- Airtable - Maintain a searchable catalog of Puma variants and product links.
- Make or Zapier - Send new dataset results into ecommerce workflows and alerts.
- Webhooks - Trigger downstream processing after an Apify run completes.
- Apify API - Retrieve run and dataset data programmatically.
Apify datasets can be exported as JSON, CSV, Excel, XML, and other formats supported by the platform.
Frequently Asked Questions
What Puma URLs are supported?
Puma category, search, tag, and country-level listing URLs are supported through start_url. Use a public listing URL that matches the products you want to collect.
Can I provide multiple Puma URLs in one run?
No. The input schema accepts one start_url per run. Run the Actor separately for additional URLs or create separate scheduled runs.
Can I scrape Puma search results?
Yes. Provide a Puma search URL, such as a URL containing a q query parameter, in start_url.
Can I export Puma data to CSV or Excel?
Yes. Apify dataset results can be downloaded as CSV, Excel, JSON, XML, and other supported formats.
Why is a field missing from one product?
Puma does not publish every field for every listing, and the Actor removes empty values. Check the product's available source data and compare several records before treating a missing field as a run problem.
Does the Actor guarantee a specific number of products?
No. results_wanted sets the maximum target. The actual count depends on products returned by the selected listing, the max_pages limit, deduplication, and available source data.
Can I run this Actor on a schedule?
Yes. Use Apify Schedules to create recurring catalog, pricing, or merchandising snapshots.
Is it legal to collect Puma listing data?
You are responsible for complying with Puma's terms, applicable laws, privacy requirements, and any restrictions related to the data you collect. Use the Actor for legitimate, responsible data collection from publicly available listings.
Related Actors
- Namshi Product Scraper - Collect product listings, prices, discounts, ratings, and stock signals from Namshi.
- AliExpress Scraper - Extract product, pricing, review, seller, and shipping data from AliExpress.
- Farfetch Pricing Scraper - Collect fashion product pricing, availability, brand, and variation data from Farfetch.
- Trendyol Product Scraper - Gather product, seller, pricing, rating, category, and image data from Trendyol search pages.
Support
For issues or feature requests, use the Issues tab on the Actor page in Apify Console. Include the input URL, run details, and an example of the missing or unexpected output when reporting a problem.
Legal Notice
This Actor is intended for legitimate collection and analysis of publicly available Puma listing data. You are responsible for complying with Puma's terms of service, applicable laws, privacy rules, and your organization's data-governance requirements.