Why copy and paste breaks PDF tables

In a scanned PDF the table is part of a page image, so there are no cells to copy. Even in some digital PDFs, the text is stored as positioned fragments rather than rows and columns. Getting a usable spreadsheet means detecting the table structure on the page.

What you get

When a table is detected, DocUnlocked exports its rows and columns to Excel (XLSX) and CSV. You also get the page text as DOCX, TXT, Markdown, HTML and JSON. When no reliable table structure is found, the spreadsheet files say so instead of guessing. Formulas and cell formatting are not reconstructed.

Try it on your own table

Pick the one or two pages that contain the table and run the free preview: no card, no account, downloads included. If the spreadsheet is usable, convert the whole document once: $0.99 up to 20 pages, $4.99 up to 100, $9.99 up to 200, up to 200 MB. No subscription.

Review cell by cell

Merged cells, faint lines and handwritten entries can shift values between columns. Check totals and key figures against the source before using the spreadsheet.

Questions before uploading.

Can I extract a table from a scanned PDF for free?

Yes, on up to two selected pages, from a source of up to 10 MB and 20 PDF pages. Downloads are included.

Does every PDF table become an Excel file?

No. XLSX and CSV contain rows only when a table is detected. Otherwise they contain a no-table notice.

Are formulas kept?

No. You get the cell values; formulas and formatting are not reconstructed.

How much does a full document cost?

$0.99 up to 20 pages, $4.99 up to 100 pages and $9.99 up to 200 pages, one-time, with no subscription.

Related conversions

See also scanned pdf to word, image to excel, scanned PDF to editable text and reviewable tables from scanned PDFs, or return to the DocUnlocked hub.

Try 2 pages free