PDF → CSV / JSON
PDF data you can
trace back.
Extract lines and columns from a PDF catalog. Every result keeps its source page and position; uncertain data stays visible for review.
01 Your document
02 Traceable results
You have not processed a document yet.
Your data will appear here
No sample results are preloaded.
Choose a real PDF to get started.
Review the output before importing it: a line may be a header, a note or part of a description. It does not automatically represent a product.
| Page | Extracted line | Fields | Warnings |
|---|
Built for agent workflows
The same engine runs in Node.js. The local API accepts a PDF and returns JSON with coordinates, original values and warnings. It does not execute instructions found in the document.
Integration and limits →What this beta does not do
It does not include OCR, guess currencies or product references, or certify an ERP import. Headers, notes and split lines remain available for review.