PDF Table Extractor - Tables to CSV (Excel-ready) avatar

PDF Table Extractor - Tables to CSV (Excel-ready)

Pricing

from $30.00 / 1,000 table page processeds

Go to Apify Store
PDF Table Extractor - Tables to CSV (Excel-ready)

PDF Table Extractor - Tables to CSV (Excel-ready)

Pull tables out of PDF reports, statements and price lists into clean CSVs that open straight in Excel. One CSV per table plus a combined dataset. Your documents, your data - nothing scraped, nothing stored.

Pricing

from $30.00 / 1,000 table page processeds

Rating

0.0

(0)

Developer

Eastwood Apps

Eastwood Apps

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

21 hours ago

Last modified

Share

Pull the tables out of PDF reports, bank statements, price lists and invoices into clean CSVs that open straight in Excel or import into any spreadsheet tool.

Quick start

  1. Paste direct links to your PDF files into PDF URLs. Use direct-download links - a Google Drive or Dropbox share page link returns a web page, not the file.
  2. Run. Each detected table becomes its own CSV in the run's key-value store (doc1-page2-table1.csv), and every row also lands in one combined dataset you can export as CSV, Excel, or JSON.

Example

A supplier price list with a 4-column table becomes:

Item,Qty,Unit,Amount
Cable 2.5mm TPS 100m drum,6,186.40,1118.40
Switchboard chassis 24 pole,1,412.75,412.75

How it works

Tables are detected from the PDF's own text geometry: lines whose entries align into shared column positions are recognised as tables, column by column, row by row. No AI guessing, no hallucinated cells - if a value is in the CSV, it was on the page at that position.

What it does

  • One CSV per detected table, named by document, page and table number
  • A combined dataset of every row (with page/table/row references) for bulk export
  • Multi-page documents and multiple tables per page handled
  • Honest flags: a document with no detectable tables is flagged review with the reason (unruled/irregular layouts and scanned PDFs are the usual causes), never padded with invented rows

What it does NOT do

  • No scraping, no logins, no third-party data. You supply your own documents, they are processed inside YOUR Apify account, and the developer never sees or stores them.
  • Scanned (image-only) PDFs are not yet supported - they are flagged, not guessed.
  • Heavily merged or borderless-irregular layouts may split imperfectly - check the review flags and open an issue with the document type; each added layout improves the product.

Pricing

Pay-per-event: a few cents per page that actually yields a table. Documents that fail or contain no detectable tables are not charged as processed. No subscription.

Terms of use

Process only documents you own or are authorised to process. Not affiliated with or endorsed by Microsoft Excel or any vendor named for compatibility.