Walmart Listings Scraper
Pricing
from $2.99 / 1,000 product listings
Walmart Listings Scraper
Extract public Walmart category and search product listings, including names, prices, sponsorship flags, brands, images, ratings, seller, availability, badges and URLs when visibly available.
Pricing
from $2.99 / 1,000 product listings
Rating
0.0
(0)
Developer
w3crawler
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Categories
Share
Extract public product listings from Walmart category and search pages for catalog analysis, price tracking, and market research. The Actor accepts searchUrls plus the compatible urls, startUrls, and url aliases, and caps output with maxItems.
Input
{"searchUrls": ["https://www.walmart.com/search?q=seed+starting+mix"],"maxItems": 20,"timeoutMs": 30000}
Only official HTTPS Walmart.com and Walmart.ca hosts are accepted. Direct HTTP is attempted first. A browser-rendered retry is used only for successful/2xx public responses that contain no parseable listings; HTTP boundaries do not trigger a needless second request. Standard Apify Proxy is used only when enabled in proxyConfiguration.
Output
Listing records
Successful listing rows include source/listing URLs, rank, capture time, item ID, name, description, category, current/original/unit price, currency, brand, images, rating, review count, seller, availability, badge, and sponsored state when literally published. Request transport and challenge details remain operational metadata.
Diagnostics
If access is blocked, the response is oversized or redirected off-platform, or no public listing can be parsed, the Actor emits an exact diagnostic row with only url, error, errorCode, and scrapedAt. Walmart’s HTTP 456 and 200-status CAPTCHA/verification responses are treated as access boundaries; HTTP status, transport evidence, and fallback counts remain in OUTPUT and logs. The Actor does not log in, solve CAPTCHAs, or perform anti-bot evasion.
You can download the dataset in various formats such as JSON, HTML, CSV, or Excel.
Extraction behavior
The parser reads JSON-LD, embedded page state, and product cards already present in returned HTML. Values are normalized, placeholder values are omitted, credentials are rejected, and duplicates are suppressed by item identity/product URL. Respect Walmart’s terms, robots guidance, rate limits, and applicable law.
Local verification
npm testnpm run lintnpm run schemaapify run --purge --input-file qa-input-279.jsonnpm run validate
Set APIFY_LOCAL_STORAGE_DIR to a batch-specific relative path for isolated QA. No Apify login, cloud push, or production promotion is performed by local tests.