
David Quiza
Building CatalogTrace: PDF-to-JSON/CSV extraction with source coordinates and review warnings for data pipelines and AI agent workflows.
catalogtrace-kiza.pages.dev/
Joined September 2026
ACTOR STATS
1 public Actor
2 total users
1 monthly user
>99% runs succeeded
Tools for traceable document data in automated workflows.
CatalogTrace extracts selectable-text PDF catalog lines and columns into JSON and CSV, retaining source page numbers, coordinates and review warnings. It supports direct public HTTPS PDF links or base64 input.
The private beta has been tested on a real supplier catalog. Results require review: a line is not automatically a product. OCR, cross-page merging and ERP writeback are not included.
Try the browser-based tool at https://catalogtrace-kiza.pages.dev/ — it processes files locally in your browser. Hosted Actor runs use Apify storage.
Store publication and live payments are not enabled yet.