PDF to Excel
Get the rows and columns of a PDF into an editable spreadsheet instead of retyping them. Everything runs in your browser — your file never leaves your device.
How the table detection works
This tool reads the text of a PDF with its position on the page, then groups it into rows and splits each row into columns wherever there is a wide gap — the way a table lays its columns out. Plain numbers become real numeric cells, so totals and formulas work straight away. It is a best-guess reconstruction, not magic: clean, well-aligned tables come out great, while merged cells, nested tables or messy spacing may need a little tidying. It all runs in your browser — your PDF is never uploaded.
Is my PDF uploaded anywhere?
No. Everything — reading the PDF, detecting columns, building the .xlsx — happens in your browser. Your data never leaves your device.
The columns didn't line up perfectly — why?
Columns are detected from the spacing between values, because a PDF does not actually store a table grid. Neatly spaced tables come out well; when values run close together or a cell is empty, the split can be off. It is meant to save you most of the retyping, then a quick tidy.
Do numbers stay as numbers?
Yes. Plain numbers (including ones written with a comma decimal) become real numeric cells, so you can sum and calculate with them immediately — no "number stored as text" warnings.
Does it work on scanned PDFs?
It will read a scan with OCR, but scans carry no column positions, so each line comes out in a single column. It works best on digital PDFs, where the real text and layout are available.
What about multiple pages?
All pages go onto one sheet, with a blank row between pages so you can see where each one starts.
Can it open a password-protected PDF?
No — remove the password in your PDF reader first, then convert the unlocked file.
Getting the numbers out of a PDF without retyping
Some of the most tedious work in any office is copying figures out of a PDF — a bank statement, a supplier invoice, a price list or a report — cell by cell into a spreadsheet. It is slow, and every manual keystroke is a chance to fat-finger a number. This tool reads the rows and columns straight out of the PDF and hands you an Excel file, so instead of an afternoon of typing you get a spreadsheet you can sort, total and chart in seconds. For anyone who works with data trapped in PDFs, that is a genuine time-saver.
How columns are worked out
A PDF does not actually store a table — it stores text at positions on the page. The tool reads each piece of text with its coordinates, groups the pieces into rows by their height on the page, and splits each row into columns wherever there is a wide gap, which is exactly how a table lays its columns out visually. That works beautifully on clean, well-aligned tables. It is a best-guess reconstruction, though, so when values sit very close together, or a cell is blank, the split can occasionally land in the wrong place.
Numbers stay numbers
A detail that makes a real difference: plain numbers come through as actual numeric cells, not text. That means you can sum a column, build a formula or make a chart the moment the file opens, with no "number stored as text" warnings to clear first. The tool also understands a comma decimal, so a European-style "149,50" becomes the value 149.50 rather than a piece of text. Dates and codes stay as text, which is usually what you want, so a reference number keeps its leading zeros instead of being mangled into a plain integer.
Tidying up, and where it shines
Think of the result as a strong first draft that saves you the typing, not a flawless import. Clean, gridded tables — invoices, statements, exports — come out ready to use. Messier sources with merged cells, nested tables or uneven spacing may need a quick tidy: nudging a value into the right column, or splitting a row that ran together. That few minutes of cleanup is still a fraction of retyping everything, and you keep the original PDF as the source of truth if you ever need to check a figure.
Scans and privacy
The tool works best on digital PDFs, where the real text and its positions are available. It will read a scanned PDF with OCR, but a scan carries no column positions, so each line comes out in a single column that you would then split yourself — useful, but not the same as a clean digital table. As with everything here, the whole process runs in your browser: your data is never uploaded, logged or stored, which matters when the file is a bank statement or a customer list.
