PDF Table Extractor
Get tables out of a PDF without retyping them. The extractor looks for blocks of lines whose text is aligned in columns, works out the columns of each table separately, shows you a preview of every table it found and exports them as CSV files and as an Excel workbook with one sheet per table. Numbers such as 1,219.99 and accounting negatives like (45.10) become real numbers.
- Files stay on your device
- No sign-up
- Free to use
How to use PDF Table Extractor
- Drop a PDF onto the page.
- Optionally limit the pages and the minimum table size.
- Click “Find tables” and check the previews.
- Download the CSV files or the Excel workbook.
PDF Table Extractor features
Table detection
Finds aligned blocks among ordinary text.
Per-table columns
Each table gets its own column positions.
Previews
First rows of every table shown before download.
CSV and Excel
One CSV per table plus an .xlsx with a sheet each.
Number conversion
Thousands separators, currency signs, (negatives).
Private
Processing happens on your device.
When to use PDF Table Extractor
- Pulling figures from financial reports.
- Reusing price lists and catalogues.
- Extracting statistics from research papers.
- Moving invoice line items into a spreadsheet.
PDF Table Extractor FAQ
How is this different from PDF to Excel?
PDF to Excel converts whole pages, every line becoming a row. The table extractor finds the tables inside the pages and exports only those, each with its own columns.
Does it work with scanned PDFs?
Only after OCR. A scan is an image without text positions; run PDF OCR first to add a text layer.
Why are some columns merged or split?
Columns are inferred from where text starts. Cells that span columns or are centred unusually can be placed in a neighbouring column; check the preview and adjust in your spreadsheet.
What does “Minimum rows” do?
It ignores small aligned blocks, such as a two-line address, that are not real tables.
Are numbers converted?
Yes, when the option is on: 1,219.99 becomes 1219.99 and (45.10) becomes −45.1 in the CSV and the workbook. The header row is never converted.
Is my PDF uploaded?
No. Everything happens in your browser.
How tables are found in a PDF
A PDF does not store tables as tables. It stores pieces of text with positions on the page, and the reader only sees a table because those pieces line up. Extracting a table therefore means reconstructing that structure from the positions.
The extractor first rebuilds the lines of each page and splits a line into cells wherever there is a wide gap between words. Consecutive lines with two or more cells that sit close together vertically form a candidate table. Within it, the left edges of all cells are clustered into columns, and columns used by only one row are ignored as wrapped text.
Each table is previewed so you can confirm what was found. The first row is treated as the header; in the following rows, numbers formatted for reading – thousands separators, currency symbols, negatives in brackets – are turned into plain numbers so that sums and charts work immediately in a spreadsheet.
CSV files are written in UTF-8 with a byte order mark so Excel opens accented characters correctly, and the workbook keeps every table on its own sheet.