HTML and PDF Table Extractor avatar

HTML and PDF Table Extractor

Pricing

$10.00 / 1,000 result delivereds

Go to Apify Store
HTML and PDF Table Extractor

HTML and PDF Table Extractor

Extract tables from public HTML or text-layer PDFs. Try is a live Wikipedia GDP table. No OCR. $0.01 / row.

Pricing

$10.00 / 1,000 result delivereds

Rating

0.0

(0)

Developer

Zentra

Zentra

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

7 days ago

Last modified

Share

World GDP 126,295,331 is a live Wikipedia cell.

Extract tables from public HTML or text-layer PDFs. Try is a live Wikipedia GDP table. No OCR. $0.01 / row.

Last run

8 row(s). First row: World GDP 126,295,331

{
"title": "World GDP 126,295,331",
"buyerName": "Wikipedia",
"sourceName": "Wikipedia GDP table",
"sourceUrl": "https://en.wikipedia.org/wiki/List_of_countries_by_GDP_(nominal)"
}

Run it

Paste this on the Input tab (JSON) and press Start. That is the same input the last-run used.

{
"sourceMode": "startUrls",
"outputMode": "buyer-ready-records",
"maxItems": 8,
"startUrls": [
{
"url": "https://en.wikipedia.org/wiki/List_of_countries_by_GDP_(nominal)"
}
]
}

What lands in the dataset

  • HTML tables + text-layer PDF
  • Live Wikipedia on Try
  • Row-level billing

Use it when

Point at a Wikipedia table or HTML page and take cells as rows. World GDP is the Try.

Researchers who need a table as JSON.

Price

$0.01 / table row — you pay $0.01 per billed result. A 10-row Try is about $0.10.

Schedule it. A one-off curiosity click is not the product.

What you should know

  • No OCR.
  • Public pages/PDFs.

FAQ

Is the last-run a fixture? No. The JSON above is from the public source in Run it.

Docling cloud? No. Public HTML or text-layer PDFs. Try is a Wikipedia GDP table.