Go to example tasks
Extract the label from a W3C PDF fixture
Created by
Automation Lab
Map the identifying label in W3C's public PDF test fixture to a typed Excel row and retain the exact source page and matched text.
PDF to Excel Extractorautomation-lab/schema-guided-pdf-to-excel-extractor
Document
Status
Pages
Extracted values
+4 fieldsTextNumberBooleanListObject
Input
PDF documents(required)
Public PDF URL:https://www.w3.org/WAI/ER/tests/xhtml/testfiles/resources/pdf/dummy.pdf
Document name:W3C PDF fixture
Field mapping(required)
Output key(required):documentLabel
Excel header:Document label
Regular expression(required):^(Dummy PDF file)$
Required:true
Only search this page:1
Output fields
Document
Status
Pages
Extracted values
Warnings
Source
Workbook key
Processed at
Sign up on Apify01
Create your Apify account to access the PDF to Excel Extractor.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
