Go to example tasks
Get the table of contents of a PDF
The bookmark outline the document carries, with each heading, its level and the page it points to. It is the cheapest way to see how a long report is organised, and to decide which page range is worth extracting in full.
PDF to JSON Extractor — Tables, Fields & Structurepower_on/pdf-to-json-extractor
PDF
Pages
Outline (heading, level, page)
Pages charged
TextNumberBooleanListObject
Input
PDF files(required):https://www.irs.gov/pub/irs-pdf/p15.pdf+1
Max pages per document:1
Output formats:json
Output fields
PDF
Pages
Outline (heading, level, page)
Pages charged
Sign up on Apify01
Create your Apify account to access the PDF to JSON Extractor — Tables, Fields & Structure.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
