Go to example tasks
Convert a PDF table to page-cited Markdown and JSON
Created by
zihang
Extract the public W3C text-based table PDF into Markdown and page-level JSON with source pages and best-effort tables. No OCR. $0.005 per successful PDF plus platform usage.
PDF Text Extractor - Markdown & Page Citationsgallant_fender/pageledger
Source PDF
Status
Pages
Markdown
+1 fieldTextNumberBooleanListObject
Input
Public PDF URLs(required):https://www.w3.org/WAI/WCAG21/working-examples/pdf-table/table.pdf
Output fields
Source PDF
Status
Pages
Markdown
Processing time (ms)
Sign up on Apify01
Create your Apify account to access the PDF Text Extractor - Markdown & Page Citations.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
