Convert from PDF
What does PDF to Excel do?
PDF to Excel extracts text from one PDF into an XLSX workbook in the browser. Extractable text, approximate row grouping and page-to-worksheet order are kept. Merged-cell structure, formulas, charts, irregular table relationships and text from scans without OCR are not reconstructed. One worksheet is generated per PDF page. Complex tables may need manual cleanup.
How to use PDF to Excel
- 1Choose one PDF
Open PDF to Excel and select a single PDF. Unlock it first if needed. Scanned pages need OCR PDF before extraction can fill cells.
- 2Run automatic text and row extraction
There is no grid-drawing step. Text is grouped into approximate rows and written to worksheet cells.
- 3Download the XLSX
One worksheet is generated per PDF page. Check columns, merged headings and numbers. Manual cleanup is common on irregular tables.
- 4Rebuild formulas in Excel
The workbook holds extracted values, not live formulas or charts. Recreate calculations in Excel if you need them.
Supported files and limits
- Output format: XLSX
- Input: PDF
- Output options: automatic text and row extraction
- one worksheet is generated per PDF page; complex tables may need manual cleanup
- Files stay on this device and are removed when you close the page.
One worksheet is generated per PDF page. Complex tables may need manual cleanup. Scanned text is omitted until you run OCR.
What the output keeps
- extractable text
- approximate row grouping
- page-to-worksheet order
Extractable text, approximate row grouping and page-to-worksheet order are retained.
What may change or be lost
- merged-cell structure
- formulas
- charts
- irregular table relationships
- text from scans without OCR
Merged-cell structure, formulas, charts, irregular table relationships and text from scans without OCR are not rebuilt.
When not to use this tool
Do not use PDF to Excel to recover a working financial model, to unmerge cells perfectly, or to read a scan that has not been through OCR. Excel to PDF is the opposite direction and rasterizes sheets instead.
If processing fails
Encrypted files must be unlocked first. If the sheet is empty, the page was likely a scan or the table was drawn as lines without extractable text — run OCR PDF and retry. Very dense pages can mis-group rows; split columns in Excel afterward.
Privacy boundary
The PDF is read in this browser. Extracted cell values are not sent to a PDF Smart Kit application server. Close the page to clear the workbook from memory.
Testing evidence
A sample test record has not been published for this tool yet. Use the limits on this page as the current capability statement.