Jarir Books & Electronics Catalog Scraper
Pricing
from $6.00 / 1,000 results
Jarir Books & Electronics Catalog Scraper
Extract public Jarir Saudi books and electronics with ISBNs, authors, SKUs, brands, specifications, prices, availability, images, and category data.
Pricing
from $6.00 / 1,000 results
Rating
0.0
(0)
Developer
coolinbex
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
4 days ago
Last modified
Categories
Share
Extract public Jarir Saudi Arabia catalog pages for market research, inventory discovery, price monitoring, and content enrichment. The Actor uses only browser-visible public pages; it does not log in, call private APIs, defeat access controls, or bypass CAPTCHA pages.
What it extracts
Each dataset record can include the product title, Jarir URL and canonical URL, category, SKU/product ID, ISBN, authors, brand, current and old price, currency, availability, description, public images, and visible specification-like label/value pairs. Books are especially useful for ISBN/author data; electronics records are useful for SKU/brand/specification/price comparisons.
Input
startUrls is a request list of public www.jarir.com category, search, or product URLs. The default uses the public electronics category because it currently exposes a product record reliably; it is intentionally limited to one page and one product without detail-page enrichment so Apify health tests complete quickly. Set scrapeDetails: true for richer ISBN, specification, and availability fields. category can be all, books, or electronics. maxItems (1–1000), maxPages (1–20), scrapeDetails, includeImages, includeSpecifications, maxConcurrency (1–2), and optional authorized proxyConfiguration control bounded usage.
Example:
{"startUrls": [{ "url": "https://www.jarir.com/sa-en/electronics.html" }],"category": "electronics","maxItems": 50,"maxPages": 2,"scrapeDetails": true}
Output and limits
Records are written incrementally to the default dataset. status is ok or partial; errorMessage is populated for a failed detail request. A SUMMARY key-value record reports pages, unique products, retries/failures, blocked pages, partial records, and source-change signals. The Actor stops at the configured item/page/request bounds, uses one-to-two browser sessions, 1.5 seconds same-domain pacing, 45-second navigation timeouts, and two retries per request. Empty or blocked runs are reported explicitly with a summary and do not fabricate products.
Jarir’s storefront is client-rendered and selectors/data shapes can change. A non-ok record or sourceChanges summary signal should be reviewed before automated downstream use. Pagination and product discovery depend on public links rendered on the supplied pages.
Runtime and cost
Browser rendering is the main cost driver. A small 10–50 product run is generally the economical starting point; detail pages roughly double page requests. Use scrapeDetails: false, low maxItems, and one concurrency for previews. Apify platform compute, proxy traffic (if selected), storage, and any applicable plan minimums are billed by Apify; this Actor does not add a separate fee. For a commercial price, a practical model is per successful product record, with a higher tier for detail enrichment and an explicit pass-through for proxy/compute costs.
Safe and compliant use
Use only public pages you are authorized to collect, honor Jarir’s terms and robots guidance, keep concurrency and page bounds low, and do not use this Actor to access accounts, personal data, checkout flows, or restricted endpoints. Validate price/availability before business decisions because storefront values change.