Free tools / PDF to Markdown
PDF to Markdown converter: tables and invoice fields
Drop a PDF and get clean Markdown, every table as CSV and, for invoices and receipts, fields such as invoice number, dates, VAT and total. It runs in your browser: the file is never uploaded.
How it works
The browser reads the text and its positions from the PDF. Layout rules turn lines into headings, lists and tables, and templates find fields by their labels in English, German, French, Spanish, Dutch, Italian, Portuguese, Polish, Czech and the Nordic languages. There is no AI model: the same file always gives the same result. A field that is not found is reported as not found.
Limits
- PDFs with a text layer only. Scanned pages need OCR, which the DocToJSON Actor does.
- The first 50 pages, files up to 30 MB. For batches, DOCX files and scheduled runs use the Actor.
- Complex layouts (merged cells, multi-column pages) can need checking.
Questions
Is my file uploaded?
No. Everything runs in your browser; nothing leaves your computer.
Can it read scanned PDFs?
Not here. Scans have no text layer; the Actor reads them with OCR.
How accurate is it?
On a synthetic test set of 44 invoices in ten languages every field was found; real documents vary more, so check important numbers. See how it works without an LLM.