Books.toscrape.com Scraper - Book Catalog Data Extractor
Pricing
from $5.00 / 1,000 results
Books.toscrape.com Scraper - Book Catalog Data Extractor
Scrape book catalog data from books.toscrape.com including titles, prices, ratings, availability, and images. Perfect for testing and learning web scraping.
Pricing
from $5.00 / 1,000 results
Rating
0.0
(0)
Developer
Archit Khurana
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Categories
Share
Books.toscrape.com Scraper
Scrape book catalog data from books.toscrape.com - a practice website designed for learning web scraping.
Features
- π Extracts complete book catalog data
- π° Captures prices, ratings, and availability
- πΌοΈ Includes book images and URLs
- π Handles pagination automatically
- π― Configurable result limits
- π Uses residential proxies for reliability
Extracted Data
Each book record includes:
- Title - Full book title
- Price - Current price
- Rating - Star rating (One to Five)
- Availability - Stock status
- Image URL - Book cover image
- Book URL - Link to book detail page
- Page Number - Which catalog page it was found on
Input Configuration
Start URL
The URL to begin scraping from.
- Default:
http://books.toscrape.com/ - Example:
http://books.toscrape.com/catalogue/category/books/travel_2/index.html
Maximum Results
Maximum number of books to scrape.
- Default: 100
- Range: 1-1000
Usage Example
{"startUrl": "http://books.toscrape.com/","maxResults": 50}
Output Example
{"title": "A Light in the Attic","price": "Β£51.77","rating": "Three","availability": "In stock","image_url": "http://books.toscrape.com/media/cache/2c/da/2cdad67c44b002e7ead0cc35693c0e8b.jpg","book_url": "http://books.toscrape.com/catalogue/a-light-in-the-attic_1000/index.html","scraped_at": "http://books.toscrape.com/","page_number": 1}
Cost
- Per Result: $0.005
- Per Run: $0.05
- Example: 100 books = $0.05 (run) + $0.50 (results) = $0.55
About books.toscrape.com
This is a sandbox website specifically created for practicing web scraping. It contains 1000 books across 50 pages with no bot protection, making it perfect for testing scrapers.
Notes
- Site has LIGHT protection (scraping-friendly)
- No authentication required
- Respects robots.txt
- Uses Apify residential proxies when available
- Suitable for learning and testing purposes
Support
For issues or questions, please open an issue on the actor's page.