HTML and PDF Table Extractor
Pricing
$10.00 / 1,000 result delivereds
HTML and PDF Table Extractor
Extract tables from public HTML or text-layer PDFs. Try is a live Wikipedia GDP table. No OCR. $0.01 / row.
Pricing
$10.00 / 1,000 result delivereds
Rating
0.0
(0)
Developer
Zentra
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
7 days ago
Last modified
Categories
Share
World GDP 126,295,331 is a live Wikipedia cell.
Extract tables from public HTML or text-layer PDFs. Try is a live Wikipedia GDP table. No OCR. $0.01 / row.
Last run
8 row(s). First row: World GDP 126,295,331
{"title": "World GDP 126,295,331","buyerName": "Wikipedia","sourceName": "Wikipedia GDP table","sourceUrl": "https://en.wikipedia.org/wiki/List_of_countries_by_GDP_(nominal)"}
Run it
Paste this on the Input tab (JSON) and press Start. That is the same input the last-run used.
{"sourceMode": "startUrls","outputMode": "buyer-ready-records","maxItems": 8,"startUrls": [{"url": "https://en.wikipedia.org/wiki/List_of_countries_by_GDP_(nominal)"}]}
What lands in the dataset
- HTML tables + text-layer PDF
- Live Wikipedia on Try
- Row-level billing
Use it when
Point at a Wikipedia table or HTML page and take cells as rows. World GDP is the Try.
Researchers who need a table as JSON.
Price
$0.01 / table row — you pay $0.01 per billed result. A 10-row Try is about $0.10.
Schedule it. A one-off curiosity click is not the product.
What you should know
- No OCR.
- Public pages/PDFs.
FAQ
Is the last-run a fixture? No. The JSON above is from the public source in Run it.
Docling cloud? No. Public HTML or text-layer PDFs. Try is a Wikipedia GDP table.