Puma Product Scraper avatar

Puma Product Scraper

Pricing

from $0.99 / 1,000 results

Go to Apify Store
Puma Product Scraper

Puma Product Scraper

Extract Puma product data at scale. Scrape prices, descriptions, images, and reviews from Puma.com. Real-time monitoring, zero blocks. Perfect for price tracking, competitive analysis, and inventory management.

Pricing

from $0.99 / 1,000 results

Rating

5.0

(3)

Developer

Shahid Irfan

Shahid Irfan

Maintained by Community

Actor stats

1

Bookmarked

12

Total users

0

Monthly active users

10 days ago

Last modified

Share

What does Puma Product Scraper do?

Puma Product Scraper is a Puma product data extractor for collecting product listings from public Puma category, search, tag, and country-level listing URLs. Enter one Puma listing URL, choose how many products and pagination pages to process, and receive structured product records with identifiers, names, brands, colors, prices, ratings, reviews, availability signals, images, sizes, promotions, and product links.

The Actor is useful for ecommerce market research, competitor price tracking, catalog monitoring, merchandising analysis, and repeatable product snapshots. Results are saved to an Apify dataset and can be reviewed, exported, scheduled, or connected to other workflows.

Why use Puma Product Scraper?

  • Product catalog research - Build a structured view of Puma products from a category or search listing.
  • Price monitoring - Compare regular, sale, promotion, and best-price values across repeated runs.
  • Merchandising analysis - Review colors, images, badges, promotions, sizes, and orderability signals.
  • Review and product intelligence - Track average ratings and review counts when Puma publishes them.
  • Country-aware collection - Use Puma URLs from supported country and language storefronts; the URL determines the listing context.
  • Automation-ready datasets - Download JSON, CSV, Excel, XML, and other formats supported by Apify, or connect the dataset to integrations and scheduled workflows.

What data can you extract from Puma?

Each dataset item represents a product variant or listing record returned from the requested Puma listing context. The Actor removes null, blank, and empty nested values, so optional fields may not appear in every item.

FieldTypeDescription
sourceStringSource label, normally puma.
source_modeStringCollection mode, such as category or search.
category_urlStringPuma URL supplied in start_url.
category_pathStringActive listing path used for the record.
category_idStringPuma category identifier when available.
category_nameStringPuma category name when available.
search_termStringSearch term detected from a search URL, when applicable.
page_offsetNumberListing offset from which the record was collected.
product_idStringProduct or variant identifier from the listing.
variant_idStringPuma variant identifier.
master_idStringParent or master product identifier.
skuStringVariant SKU or product identifier used by the listing.
eanStringEAN identifier when available.
product_urlStringDirect Puma product page URL.
nameStringProduct name.
headerStringProduct header text when available.
sub_headerStringProduct sub-header text when available.
brandStringBrand label returned for the product.
color_nameStringHuman-readable variant color name.
color_codeStringPuma color value or code.
orderableBooleanWhether the returned variant is marked orderable.
orderable_color_countNumberNumber of orderable colors when available.
app_exclusiveBooleanApp-exclusive indicator when supplied by Puma.
app_only_fromStringApp-only start timestamp when available.
app_only_toStringApp-only end timestamp when available.
price_regularNumberRegular product price.
price_saleNumberSale price when available.
price_promotionNumberPromotion price when available.
price_bestNumberBest price value when available.
price_taxNumberTax amount when available.
price_tax_rateNumberTax rate when available.
is_sale_price_elapsedBooleanWhether Puma marks the sale price as elapsed.
ratingNumberAverage product rating.
review_countNumberNumber of product reviews.
score_ratingNumberAdditional score rating when available.
score_amountNumberAdditional score amount when available.
badge_labelsArrayProduct badge labels.
variant_promotion_messagesArrayPromotion messages attached to the variant.
master_promotion_messagesArrayPromotion messages attached to the parent product.
image_urlStringPrimary product image URL.
image_vertical_urlStringVertical version of the primary image when available.
image_altStringAlternative text for the primary image.
all_image_urlsArrayAvailable image URLs for the variant.
sizesArraySize groups with labels, availability, product IDs, and maximum orderable quantities.
color_optionsArrayOther color options with names, values, and image URLs.
scraped_atStringISO timestamp for when the record was collected.

How to scrape Puma product data

  1. Open Puma Product Scraper on Apify.
  2. Enter a public Puma category, search, tag, or country-level listing URL in start_url.
  3. Set results_wanted and max_pages for the size of the collection.
  4. Optionally configure proxyConfiguration if the source connection is rate-limited.
  5. Start the run and review the dataset preview.
  6. Export the dataset or connect it to your ecommerce research and monitoring workflow.

The Actor accepts one starting URL per run. It follows listing pagination up to the configured page limit and stops when the requested result count is reached, no more products are returned, or the page limit is reached.

Input Parameters

ParameterTypeRequiredDefaultDescription
start_urlStringNohttps://us.puma.com/us/en/men/shoesPuma category, search, tag, or country-level listing URL. The URL also determines the country and language context when those are included in the URL.
results_wantedIntegerNo20Maximum number of product records to save. Must be at least 1.
max_pagesIntegerNo1Maximum number of listing pagination pages to process. Must be at least 1.
proxyConfigurationObjectNo{"useApifyProxy": false}Optional Apify Proxy configuration for connections that are rate-limited.

The actor input schema does not provide a multiple-URL array, keyword field, price filter, brand filter, or configurable page-size field. Use the appropriate Puma listing URL to express the category or search target.

Usage Examples

Basic Category Extraction

Collect up to 20 products from Puma's men's shoes category:

{
"start_url": "https://us.puma.com/us/en/men/shoes",
"results_wanted": 20
}

Search URL Collection

Collect products returned by a Puma search URL. Search parameters remain part of the URL:

{
"start_url": "https://us.puma.com/us/en/search?q=running",
"results_wanted": 50,
"max_pages": 3
}

Larger Run with Apify Proxy

Process more listing pages and enable the optional proxy configuration when repeated requests need additional connection support:

{
"start_url": "https://de.puma.com/de/de/herren/schuhe",
"results_wanted": 100,
"max_pages": 5,
"proxyConfiguration": {
"useApifyProxy": true
}
}

Sample Output

This illustrative dataset item shows the main fields returned for a Puma product variant. Optional fields are omitted when Puma does not provide them:

{
"source": "puma",
"source_mode": "category",
"category_url": "https://us.puma.com/us/en/men/shoes",
"category_path": "/men/shoes",
"category_id": "mens-shoes",
"category_name": "Men's Shoes and Sneakers",
"page_offset": 0,
"product_id": "198556306179",
"variant_id": "198556306179",
"master_id": "308762",
"sku": "198556306179",
"product_url": "https://us.puma.com/us/en/pd/scuderia-ferrari-trinity-2-mens-sneakers/308762?swatch=07",
"name": "Scuderia Ferrari Trinity 2 Men's Sneakers",
"brand": "Ferrari",
"color_name": "PUMA Black-Speed Yellow",
"color_code": "07",
"orderable": true,
"price_regular": 103,
"price_sale": 103,
"rating": 0,
"review_count": 0,
"badge_labels": [
"New"
],
"image_url": "https://images.puma.com/image/upload/f_auto,q_auto,b_rgb:fafafa,w_2000,h_2000/global/308762/07/sv01/fnd/PNA/fmt/png/Scuderia-Ferrari-Trinity-2-Men's-Sneakers",
"scraped_at": "2026-04-01T10:55:22.102Z"
}

Data-quality behavior and limitations

  • Empty values are omitted - Null, blank, and empty nested values are removed from each dataset item. A missing key generally means the source did not provide a value for that product.
  • Pricing and review recovery - When listing data contains missing or zero price or review values, the Actor attempts to recover them from available product-level data. Recovery is best effort and the original value may remain missing or zero.
  • Variant-level records - The dataset is based on listing variants. Similar products in different colors or variants can therefore appear as separate records when Puma returns them separately.
  • Duplicate control - Records are deduplicated using variant or product identifiers during a run.
  • Result counts are maximums - results_wanted is a cap, not a guarantee. The final dataset can contain fewer records if the listing has fewer products, pagination ends, records cannot be mapped, or max_pages is reached.
  • Source-dependent fields - Ratings, reviews, prices, promotions, sizes, images, and orderability can vary by country, language, product, and storefront state.
  • Listing-focused collection - The supported workflow starts from Puma listing URLs. The input does not accept a list of product detail URLs or provide separate detail-page filters.
  • Public-page changes - Puma can change listing availability, URLs, fields, or storefront behavior. Check a small run after changing URL patterns and report persistent issues through the Actor page.

Tips for Best Results

  • Use a complete Puma listing URL that already represents the category or search you want to analyze.
  • Start with results_wanted: 20 and max_pages: 1 to validate the dataset before increasing the run size.
  • Use max_pages to control how far pagination proceeds; the Actor uses a fixed internal page request size.
  • For country-specific research, use the relevant Puma country and language URL rather than relying on a different storefront.
  • Treat absent optional fields as source-data limitations instead of assuming the entire record failed.
  • For recurring price or catalog monitoring, compare datasets from scheduled runs using stable identifiers such as product_id, variant_id, master_id, and sku.

Integrations and Export Formats

  • Google Sheets - Review product prices, ratings, and assortment changes in a spreadsheet.
  • Airtable - Maintain a searchable catalog of Puma variants and product links.
  • Make or Zapier - Send new dataset results into ecommerce workflows and alerts.
  • Webhooks - Trigger downstream processing after an Apify run completes.
  • Apify API - Retrieve run and dataset data programmatically.

Apify datasets can be exported as JSON, CSV, Excel, XML, and other formats supported by the platform.

Frequently Asked Questions

What Puma URLs are supported?

Puma category, search, tag, and country-level listing URLs are supported through start_url. Use a public listing URL that matches the products you want to collect.

Can I provide multiple Puma URLs in one run?

No. The input schema accepts one start_url per run. Run the Actor separately for additional URLs or create separate scheduled runs.

Can I scrape Puma search results?

Yes. Provide a Puma search URL, such as a URL containing a q query parameter, in start_url.

Can I export Puma data to CSV or Excel?

Yes. Apify dataset results can be downloaded as CSV, Excel, JSON, XML, and other supported formats.

Why is a field missing from one product?

Puma does not publish every field for every listing, and the Actor removes empty values. Check the product's available source data and compare several records before treating a missing field as a run problem.

Does the Actor guarantee a specific number of products?

No. results_wanted sets the maximum target. The actual count depends on products returned by the selected listing, the max_pages limit, deduplication, and available source data.

Can I run this Actor on a schedule?

Yes. Use Apify Schedules to create recurring catalog, pricing, or merchandising snapshots.

You are responsible for complying with Puma's terms, applicable laws, privacy requirements, and any restrictions related to the data you collect. Use the Actor for legitimate, responsible data collection from publicly available listings.

  • Namshi Product Scraper - Collect product listings, prices, discounts, ratings, and stock signals from Namshi.
  • AliExpress Scraper - Extract product, pricing, review, seller, and shipping data from AliExpress.
  • Farfetch Pricing Scraper - Collect fashion product pricing, availability, brand, and variation data from Farfetch.
  • Trendyol Product Scraper - Gather product, seller, pricing, rating, category, and image data from Trendyol search pages.

Support

For issues or feature requests, use the Issues tab on the Actor page in Apify Console. Include the input URL, run details, and an example of the missing or unexpected output when reporting a problem.

This Actor is intended for legitimate collection and analysis of publicly available Puma listing data. You are responsible for complying with Puma's terms of service, applicable laws, privacy rules, and your organization's data-governance requirements.