PDF Table Extractor avatar

PDF Table Extractor

Pricing

from $43.00 / 1,000 enriched records

Go to Apify Store
PDF Table Extractor

PDF Table Extractor

Extract tables and text from a public PDF URL. Try is a W3C table PDF. Scanned OCR is out of scope. $0.03 / PDF.

Pricing

from $43.00 / 1,000 enriched records

Rating

0.0

(0)

Developer

Zentra

Zentra

Maintained by Community

Actor stats

0

Bookmarked

1

Total users

0

Monthly active users

6 days ago

Last modified

Share

Text-layer PDFs. We do not pretend to OCR a scan.

Extract tables and text from a public PDF URL. Try is a W3C table PDF. Scanned OCR is out of scope. $0.03 / PDF.

Last run

1 row(s). First row: table.pdf

{
"title": "table.pdf",
"sourceUrl": "https://www.w3.org/WAI/WCAG21/working-examples/pdf-table/table.pdf"
}

Run it

Paste this on the Input tab (JSON) and press Start. That is the same input the last-run used.

{
"sourceMode": "startUrls",
"outputMode": "production-records",
"maxItems": 5,
"startUrls": [
{
"url": "https://www.w3.org/WAI/WCAG21/working-examples/pdf-table/table.pdf"
}
]
}

What lands in the dataset

  • Text-layer table extract
  • W3C table.pdf on Try
  • One billed document

Use it when

Give a text-layer PDF URL (W3C table.pdf on Try). Scans will not OCR themselves.

Analysts with born-digital PDFs, not a pile of scans.

Price

$0.03 / PDF — you pay $0.03 per billed result. A 10-row Try is about $0.30.

Schedule it. A one-off curiosity click is not the product.

What you should know

  • No OCR.
  • Public URL only.

FAQ

Is the last-run a fixture? No. The JSON above is from the public source in Run it.

Scanned OCR? Out of scope. Text-layer tables from a public PDF URL.