Find a tool
Tutorials

How to Convert PDF to Excel (Tables That Actually Line Up)

How to Convert PDF to Excel (Tables That Actually Line Up)

You’ve got a PDF with a table in it — a bank statement, an invoice, a price list, a report someone exported without thinking about the poor soul who’d need the numbers back — and you need it in Excel so you can sort, total and actually work with it. Retyping is out of the question. The good news is that a clean, text-based table converts to a spreadsheet in seconds. The catch is that not every PDF is clean or text-based, and knowing the difference is the whole game. Here’s how to convert PDF to Excel well, and how to fix the cases that don’t convert on the first try.

ItemQtyPricePaper104.20Toner238.00Folders250.90PDF tableextractXLSX1234ABCPaper104.20Toner238.00Folders250.90Excel spreadsheet
A text-based table in a PDF maps cell-for-cell into a real Excel sheet — rows become rows, columns become columns, and the numbers land in cells you can sort and total.

Why PDF-to-Excel is trickier than it looks

A spreadsheet knows it has rows and columns; a PDF doesn’t. Under the hood a PDF just says “put this text at this position on the page.” A table only looks like a grid because the numbers happen to line up — there’s usually no stored structure saying “this is column three.” So every PDF-to-Excel converter is really reconstructing the table: reading the position of each piece of text and inferring where the rows and columns must be. When a table is neat and text-based, that inference is accurate. When cells are merged, borders are missing, or a value wraps onto two lines, the converter has to guess — which is why results range from perfect to “everything’s in column A.” None of this is a bug; it’s the nature of the format.

How to convert a PDF to Excel with Ream

PDF to Excel reads the tables in your document and rebuilds them as spreadsheet cells:

  1. Add your PDF. Open PDF to Excel and drop in the file.
  2. Convert. The tool finds the tabular data and maps it into rows and columns.
  3. Download the spreadsheet and open it in Excel, Numbers, LibreOffice or Google Sheets.

It works best on real, text-based tables — the kind exported from accounting software, a database, or another spreadsheet. The more a table looks like a proper grid, the cleaner the result. Very complex layouts, or scans, need the extra step below.

TEXTtext-based PDFextract directlyXLSXclean cellsSCANscanned image PDFOCR firstOCRXLSXcells, after OCRText-based → direct.Scanned → OCR, then convert.
The make-or-break difference: a text-based PDF has a real text layer and converts straight to Excel. A scanned page is just an image — it needs OCR first to turn the picture of numbers into actual numbers.

Text-based vs. scanned: the make-or-break difference

Here’s the single biggest factor in whether your conversion works. A text-based PDF was created digitally — you can select the text with your cursor. A scanned PDF is a photo of a page: the “numbers” are just pixels, with no text underneath. A converter can’t extract data that isn’t there, so a scan comes out empty or garbled. The fix is OCR (optical character recognition), which reads the image and adds a real text layer. Run the scan through OCR PDF first, then convert the result to Excel. Quick test: if you can’t select the numbers with your mouse in a PDF viewer, it’s a scan and needs OCR. Our guide on making a scanned PDF searchable walks through it.

Other ways to get a PDF table into a spreadsheet

Depending on the file and the tools you have, a few routes exist:

  • Excel’s built-in import: in Excel, Data → Get Data → From File → From PDF lets you pick a table from the PDF and load it. It’s handy for text-based files, though the preview can be fiddly with complex layouts.
  • Copy and paste: for a small, simple table you can sometimes select it in a PDF viewer and paste into Excel — then use Data → Text to Columns to split it up. Fast for a few rows, painful for many.
  • Google Sheets: there’s no direct PDF import, so you’d convert to Excel first (with PDF to Excel) and then open that file in Sheets.
  • A converter like PDF to Excel: the simplest path when you just want the whole table out cleanly, with no software to install.
Method Best for Watch out for
PDF to Excel Whole tables, no setup Needs a text-based PDF (OCR scans first)
Excel → Get Data → From PDF Picking one table from a file Preview fiddly on complex layouts
Copy & paste + Text to Columns A few simple rows Tedious; breaks on merged cells
OCR, then convert Scanned statements & receipts Verify numbers after OCR

Getting a clean result

A few habits dramatically improve the output:

  • Check the numbers, not just the layout. After any conversion — especially from a scan — spot-check a totals row. OCR can read an 8 as a 3 or lose a decimal point, and one wrong digit in a spreadsheet is worse than an obvious mess.
  • Mind the locale. “1.234,50” (European) versus “1,234.50” (US) trips up imports. If numbers land as text or with the wrong separators, fix the decimal/thousands settings on import rather than editing cells by hand.
  • Expect merged cells to need a nudge. Headers that span several columns, or a value that wraps to two lines, are where converters guess. Fixing a couple of cells afterwards is normal.
  • One table per sheet. If a page has several tables, converting them into separate sheets (or splitting the pages first) keeps each grid clean.
  • Just need the words, not a grid? If the document is prose rather than tabular, PDF to Word is the better target — see what survives a PDF-to-Word conversion.

Is it private, and is it free?

PDF to Excel is free with no sign-up. Rebuilding tables is heavier work than a simple page edit, so it runs on our server (marked with a “Server” badge): the file is sent over an encrypted HTTPS connection to a hardened, network-isolated sandbox, converted, and deleted immediately afterwards — never stored or shared. If your document is a scan, the OCR step runs the same way. For financial data you’d rather not send anywhere, that honest boundary is worth knowing up front.

Frequently asked questions

Can I convert a scanned PDF to Excel? Not directly — a scan is an image with no text to extract. Run it through OCR first to add a text layer, then convert the result.

Why are all my columns squashed into one? The converter couldn’t tell where the columns break — usually because borders are missing or cells are merged. A cleaner, text-based source helps; otherwise fix the split with Text to Columns in Excel.

How do I convert a PDF to Excel for free? Add the file to PDF to Excel, convert, and download the spreadsheet — free, no account, no watermark.

Can I convert a PDF straight into Google Sheets? Sheets has no direct PDF import. Convert to Excel first, then open the .xlsx in Google Sheets.

Will formulas come across? No — a PDF stores results, not formulas, so you get the values. You can re-add formulas in Excel once the numbers are in cells.

The numbers imported as text — why? Usually a locale mismatch (decimal vs. thousands separators) or leftover spaces. Set the right number format on import, and the values will calculate normally.

Is my financial data safe? The conversion runs in an isolated sandbox over HTTPS and the file is deleted right after — nothing is stored. Still, for the most sensitive files, keep in mind this is one of the server-side tools.

Convert a text-based table and you’ll have working cells in seconds; hit a scan and the answer is simply “OCR first, then convert.” Either way, check a total or two before you trust the sheet — and if the file is really prose with a table bolted on, PDF to Word may be the tool you actually want.