Extract tables from PDF
Every table in your PDF is detected automatically, including borderless tables, merged cells and tables that continue over several pages. Check each one in the preview, then download what you need.
What you get
- A list of all tables found, with captions
- A spreadsheet-style preview of each table
- Download per table (CSV or Excel), or all tables at once
- Warnings when a table looks incomplete
How it works
- Tables are located on every page by a layout model, with or without ruling lines.
- Rows, columns, merged cells and multi-level headers are reconstructed by a table-structure model.
- A table that continues on the next page (with a repeated header) is joined back into one table.
Good to know
- Tables from scanned pages depend on scan quality. Always compare important numbers with the original.
- Very dense scanned tables (old statistical reports, for example) are the hardest case. If rows look merged, you'll see a warning.
- Nothing is invented: a cell that cannot be read stays empty.
Questions
Do I have to select the table area?
No. Tables are found automatically on every page.
What about tables split over two pages?
If the continuation has the same columns (usually with a repeated header), it is joined into one table and the repeated header is removed.
CSV or Excel?
CSV is best for scripts and databases. Excel keeps numbers and dates as real values and puts each table on its own sheet.