PDF to Excel
Extract table-like data from a PDF into an Excel (.xlsx) file β processed entirely in your browser, nothing is uploaded.
Free PDF to Excel Converter β Extract Table Data in Your Browser
Pulling data out of a PDF and into a spreadsheet is one of the more frustrating everyday document tasks β a bank statement, a price list, or a report table that you need to actually calculate with, sort, or filter, rather than just read. Because a PDF has no real concept of "rows" and "columns" the way a spreadsheet does β it only knows where each piece of text sits on the page β converting one back into a proper table means reconstructing that structure from position alone.
How This Tool Works (and Its Honest Limits)
This converter reads every piece of text on each page along with its exact position, then groups text that sits at roughly the same vertical position into a row, and groups text that lines up at similar horizontal positions into columns β essentially reverse-engineering the table grid from where the text visually sits. This approach works well for PDFs that were generated from an actual spreadsheet or a cleanly formatted table, since the text in those documents tends to line up predictably. It will not work for a scanned PDF (a photograph or scan of a printed page treated as one big image) since there's no actual text to read in that case β only pixels. It also struggles with tables that have merged cells, multi-line wrapped text inside a single cell, or layouts where columns aren't consistently aligned, since the tool has no way to know your intended table structure beyond text position.
Using the Column Sensitivity Setting
The "Column Sensitivity" option controls how large a horizontal gap between pieces of text needs to be before the tool treats them as separate columns rather than the same one. If your source table has narrow, tightly-packed columns, choose "Tight" so nearby columns don't get merged together. If your table has generously spaced-out columns, choose "Loose" so a single column's text doesn't get mistakenly split into two. "Medium" is a reasonable starting point for most standard tables β if your result looks off, try re-converting with a different sensitivity setting.
Step-by-Step Usage
Upload your PDF and the tool will detect its total page count. Choose a column sensitivity setting that roughly matches your table's spacing, then click "Convert to Excel." Each page of the PDF becomes its own sheet in the resulting Excel file (labeled "Page 1," "Page 2," and so on), so a multi-page PDF produces a multi-sheet workbook you can review and clean up as needed.
Tips for Best Results
Always review the extracted spreadsheet before relying on it β because this is a best-effort text-position reconstruction rather than true table recognition, some rows or columns may need manual adjustment, particularly around merged headers or wrapped text. If your source PDF was originally created from Excel, Word, or Google Sheets (rather than scanned), you'll typically get the cleanest results, since that text tends to be precisely aligned. For scanned documents, you'll need an OCR (optical character recognition) tool first to turn the scan into real text before a position-based extraction like this one can work at all.
Privacy and Cost
Every step of the extraction and Excel file creation happens locally in your browser using JavaScript libraries β your PDF is never uploaded anywhere. There's no sign-up, no watermark, and no daily usage limit.
Frequently Asked Questions
Why are some of my table's columns merged together in the result?
This usually means the "Column Sensitivity" setting is too loose for your table's actual spacing β text that should be in separate columns is close enough together that the tool grouped it into one. Try re-converting with the "Tight" sensitivity setting.
Why is my table split into extra, unwanted columns?
The opposite problem: the sensitivity setting is too tight, so text within a single column is being treated as multiple columns. Try "Loose" sensitivity and re-convert.
Can this tool extract text from a scanned PDF?
No β a scanned PDF is essentially a photograph of a page with no embedded, readable text, so there's nothing for this tool to detect positions for. You'd need to run the scan through an OCR tool first to convert it into real text before any position-based table extraction is possible.
Will formulas or formatting from the original document carry over?
No β a PDF doesn't retain spreadsheet formulas or cell formatting to begin with (a PDF only stores how the final result looked as flat text and lines), so the Excel output will contain the extracted text values only, without any original formulas, colors, or number formatting.