1. Your PDFs
Invoices, statements, payslips, reports, purchase orders: PDFs with selectable text (exported from software, downloaded from a portal). A scanned PDF with no text layer is flagged.
| # | File | Pages | Status |
|---|
2. Fields to extract
Each field becomes a column. Three methods: the text after a label ("Total", "Invoice No"), a regular expression, or a ready-made detector (date, total amount, VAT number, IBAN…).
| Column name | Method | Setting | Type | If several |
|---|
3. Result
Empty cells are orange. Every cell can be fixed directly in the table. Click a row to see the text read from that PDF, with extracted values highlighted.
Frequently asked questions
My invoices come from different suppliers with different layouts. Will it work?
Yes for common fields. The detectors (invoice number, date, total, net, VAT, VAT ID, IBAN) look for the words around a value, not its position on the page. For an unusual supplier, add a label or a regular expression and save it as a template.
It says "scanned PDF". What now?
The PDF is an image without selectable text, and the tool does not do OCR. Run it through OCR first (for example Acrobat's Recognize Text, or a scanner that makes searchable PDFs), then drop it again.
Why not use Power Query in Excel?
Power Query's PDF connector imports the tables it finds in each file. When layouts differ, the tables change from file to file and the query breaks or mixes columns. Here each field is found by its label or format in every PDF.
Do dates and amounts arrive as numbers in Excel?
Yes. In the .xlsx file dates are real dates and amounts real numbers, so you can sort, filter and sum. Formats such as 1,234.56 and 1.234,56 are recognised.
What does it cost?
The full table is shown before you pay, and export is free for up to 5 PDFs at a time. Beyond that, $4.99 for 24 hours or