Extract PDF tables into Excel without losing track of columns
A PDF stores a page layout, not a spreadsheet. Values that look aligned may be separate text fragments. Utilgo combines table lines and text alignment to place them into cells. Detected tables receive separate sheets; readable page text is retained on reference sheets. The Report sheet records the extraction method and cases that need review.
Step by step
- Upload a PDF and optionally select pages such as 1,3-5. A conversion accepts up to 30 pages.
- Start with automatic detection. If columns are misplaced, try ruled tables, borderless tables or column spacing.
- For scans, choose the document language. At most 10 pages can use OCR. Download the workbook and compare Report with the original.
Example scenario
Example: a price list contains product code 00123 and amount 1,234.56. Text mode preserves both spellings. Choosing dot-decimal numeric conversion makes suitable amounts usable in calculations, but numeric-looking codes can also change. For mixed identifiers and amounts, keep text and convert only the amount column in Excel.
Check your result
Check the first and last row, subtotals and page boundaries. Merged headers, wrapped cells and crooked scans can split incorrectly. OCR may confuse 0 with O or 1 with I. Reference text does not prove every table was detected. Formulas and workbook relationships cannot be recovered from a PDF.
Frequently asked questions
What if no table is detected?
Readable content remains in a single-column reference sheet, with a note in Report. It is not labelled as a successfully extracted table.
Can PDF content become an Excel formula?
No. Imported content is not executed as a formula. Cells remain text unless numeric conversion is selected.
Keep an original copy before processing your file. Examples are illustrative; results depend on your document.