Convert PDF to Excel Free (Tables That Actually Work)
Convert PDF to Excel free and get real rows and columns, not a mess. Text PDFs vs scanned statements, OCR, and post-conversion cleanup tips.

You have a table trapped in a PDF — a bank statement, an invoice, a price list, a 30-page report appendix — and retyping it would take an hour and introduce typos. Converting PDF to Excel free takes about 30 seconds and gets you real rows and columns you can sort, filter and sum. The catch: PDF is a genuinely hostile format for tables, so the right workflow depends on what kind of PDF you're holding. Here's the whole picture.
Why PDFs are so bad at tables
A PDF doesn't contain a table. It contains instructions like "draw '1,240.50' at x=412, y=305" — hundreds of positioned text fragments plus some lines. The rows, columns and cells you see exist only in your head. There is no table structure in the file to copy out, which is why selecting a table in a PDF viewer and pasting into Excel dumps everything into one column.
A converter has to reverse-engineer the grid: cluster fragments into rows by their y-coordinates, detect column boundaries from x-positions and ruled lines, and rebuild cells. Modern extraction does this well on clean tables — but it's inference, not copying, so a quick check of the output against the original is always part of the job.
Convert a PDF table to Excel in the browser
Open the converter
Go to PDF to Excel in any browser — no install, no account.
Upload the PDF
Free up to 20 MB and 50 pages — a year of monthly statements fits comfortably.
Convert
Tables are detected and rebuilt as spreadsheet cells. 10-40 seconds for a typical statement.
Download the .xlsx
Opens in Excel, Google Sheets or LibreOffice. Files are auto-deleted from our server after 1 hour.
Sanity-check the numbers
Sum a column you know the total of and compare against the PDF. If it matches, the extraction is trustworthy.
The free plan covers 10 operations a day; Pro ($9/mo) removes all limits if you process statements in bulk.
Text PDF vs scanned PDF — the 5-second test
Everything downstream depends on which kind you have. Open the PDF and try to select the text with your cursor:
- Text selects → born-digital PDF (exported from software). Convert directly; expect near-perfect numbers.
- Nothing selects → scanned PDF (a photo of paper). There is literally no text in the file yet — convert it directly and you'll get an empty or garbage sheet.
For scans — which is what most bank statements printed and re-scanned, receipts and old records are — the pipeline is:
- Run the file through OCR PDF to recognize the printed characters and add a real text layer.
- Feed the OCR'd PDF into PDF to Excel.
- Verify totals. OCR on clean print runs 95-99% accurate, and its classic misreads are exactly the ones that hurt in finance: 0↔8, 1↔7, 5↔6. Checking that column sums match the statement's own printed totals catches nearly all of them.
Cleaning up the spreadsheet after conversion
Even a good conversion leaves fingerprints. The usual suspects, with the fast fixes:
| Problem | What you see | Fix |
|---|---|---|
| Merged cells | Multi-line PDF cells become merged blocks that break sorting | Select all → Home → Merge & Center → Unmerge Cells, then fill gaps |
| Numbers stored as text | Left-aligned numbers, SUM returns 0 | Select column → Data → Text to Columns → Finish forces re-parse |
| Currency symbols glued on | '$1,240.50' as one text string | Find & Replace the symbol with nothing, then reformat as Currency |
| Negative amounts as (123.45) | Parentheses read as text | Text to Columns re-parses them as negatives in most locales |
| Repeated page headers | Column titles reappear every 40 rows | Filter the header text in one column → delete visible rows |
| Split or shifted columns | One PDF column became two, or rows drifted | Usually a complex layout — re-check against the original, merge manually |
Two of these are worth memorizing because they apply everywhere: Text to Columns (Data tab) is Excel's universal "re-read this column properly" button, and unmerge-then-fill turns a pretty-but-useless layout into a sortable dataset.
Alternative: Excel's own "Get Data from PDF" (Windows)
If you have Microsoft 365 on Windows, Excel has a built-in importer: Data → Get Data → From File → From PDF. It lists every table Power Query detects in the file; you pick one, optionally clean it in the Power Query editor, and load it to a sheet.
Where it shines: recurring imports. Set the query up once for your monthly statement, and next month it's Data → Refresh on the new file. Where it doesn't: it's Windows-only (the Mac version has long lagged on this feature), it can't read scans at all — no OCR — and its table detection misses tables that span pages more often than dedicated converters do. For a one-off extraction, or on a Mac or Chromebook, the browser tool is the shorter path.
Frequently asked questions
Is PDF to Excel really free?
On pdfty, yes — free up to 20 MB and 50 pages per file, 10 operations a day, no sign-up, no watermarks. Pro ($9/mo) removes the limits. Your files are deleted from the server after 1 hour either way.
Why does my converted spreadsheet show nothing, or gibberish?
Your PDF is almost certainly a scan — a photo of a page with no text layer. Run it through OCR PDF first, then convert the result. If text does select in the PDF but the output is still messy, the layout is complex — try converting just the pages with the table.
Can I convert a scanned bank statement to Excel?
Yes, in two steps: OCR PDF first, then PDF to Excel. Then verify — sum each amount column and compare with the totals printed on the statement itself.
Why won't Excel sum the numbers from my converted PDF?
They were imported as text. Select the column, then Data → Text to Columns → Finish — this forces Excel to re-parse the values as numbers. If currency symbols are glued to the values, Find & Replace them away first.
Do multi-page tables come out as one table?
Tables that continue across pages are extracted page by page; you may see the header row repeated at each page break. Filter for the header text and delete those rows — 20 seconds — and the data stacks into one continuous table.
Will formulas from the original spreadsheet come back?
No. A PDF stores only the printed values — the formulas were lost the moment the file was exported to PDF. You get the numbers; rebuild the few formulas you need. (Going the other way? See how to convert Excel to PDF without cutting off columns.)
Is uploading a bank statement safe?
On pdfty, files transfer over HTTPS, are never watermarked or shared, and are automatically deleted from the server after 1 hour. No account means no stored history tied to you. For documents you can't upload anywhere, Excel's built-in Get Data From PDF (Windows) keeps everything local.
Can I just copy-paste the table from the PDF instead?
You can try — and for a small, simple table it occasionally works. But because the PDF has no real table structure, pasting usually collapses everything into one column. For anything bigger than a few rows, a converter that reconstructs the grid saves you the manual re-splitting.
The pdfty team builds privacy-first online PDF tools — compress, convert, OCR, sign and protect. Files are deleted within 1 hour. About us →


