How to Extract Tables from a PDF into Excel
Stop retyping numbers from PDF reports. How PDF-to-Excel conversion works, when it's reliable, and how to check the result.
The slowest way to get numbers out of a PDF report is retyping them — and it's also the way most likely to introduce a silent typo into your spreadsheet. Converting the PDF to Excel extracts the tables into cells you can actually sort, filter, and run formulas on.
Where this comes up
- Bank and card statements that only download as PDF, needed in a budget sheet.
- Supplier price lists and invoices that have to land in an ERP or comparison sheet.
- Published reports and government data tables you want to analyze.
How to convert
- Open the PDF to Excel tool.
- Upload the PDF — tables are detected and mapped into rows and columns.
- Download the XLSX and open it in Excel or Google Sheets.
What converts well (and what doesn't)
A PDF has no real concept of a "table" — just text positioned on a page — so conversion is reconstruction. Clean, gridded tables exported from software convert very reliably. Merged header cells, multi-line cells, and tables that span pages are harder and may need a quick manual tidy-up. Scanned PDFs are the tough case: if you can't select the text in the PDF, it's an image, and the numbers first have to be recognized rather than extracted.
Always verify the totals
After any conversion, spot-check a few rows against the original and — if the table has a totals row — recompute it with a quick SUM. A column that sums to the printed total is strong evidence every value above it came across intact. It's a thirty-second check that catches the rare misread digit before it reaches a report.
Last updated
July 21, 2026