Skip to content
📈

PDF to Excel

Pull PDF tables into XLSX spreadsheets.

✓ Free 🔒 Secure ⚡ Fast 📁 Up to 25MB
Server processed Processed on our servers over HTTPS, then deleted within 1 hour. Files are processed automatically. No human reviews your documents, and PDFRun does not use them to train models.
Drop your file here

or click to browse — supports PDF files up to 25MB

✓ PDF to Excel — done

Processed successfully. Download below.

How to use

  1. 1 Drop or click to upload your file
  2. 2 Adjust options if shown
  3. 3 Click PDF to Excel
  4. 4 Download your result instantly
🚀 Go Pro
  • Files up to 1 GB
  • Unlimited jobs/hour
  • Batch processing (up to 100 files)
  • Priority support
Upgrade to Pro
🔒 Privacy

Files are processed securely and deleted within 1 hour. Files are processed automatically. No human reviews your documents, and PDFRun does not use them to train models.

Related Tools

Why this works

Pull tables out of a PDF into an editable Excel spreadsheet (.xlsx). Works on both born-digital PDFs (clean tables) and scanned PDFs (auto-OCR’d into structured cells).

PDFs are where data goes to die. Every bank statement, every utility bill, every research dataset arrives as a PDF whose tables you can see but can’t actually use — you can’t total a column, sort by date, or pivot by category until the numbers are in a real spreadsheet.

PDF to Excel extracts tabular data from your PDF and rebuilds it as a real Excel spreadsheet. Each detected table becomes a worksheet in the output .xlsx file. Numbers come through as numbers (not text); dates come through as dates; column headers carry over; rows preserve their grouping. For multi-table PDFs (bank statement with multiple sections, financial report with summary + detail tables), each table becomes its own sheet, named for the table’s position in the source document.

The converter handles two distinct cases. Born-digital PDFs (anything exported from Excel, accounting software, or a database) extract cleanly because the table structure was encoded into the PDF when it was created. Expect near-perfect column alignment and zero text errors. Scanned PDFs (bank statements that arrive as image-only PDFs, photographed receipts) run through OCR first, then table detection. Accuracy is high for clean modern scans but may need a manual touch-up in Excel for rows where column boundaries weren’t crisp in the source.

What extracts well: simple grid tables, multi-column financial statements, transaction lists, contact lists, schedules, price lists. What needs a touch-up: heavily merged-cell layouts (org charts in table form), tables with nested sub-tables, free-flowing text formatted to look like a table.

Numbers and dates are typed appropriately in the output — a column of dollar amounts comes through as proper numbers you can SUM(), not as text strings. Date columns are recognised and formatted as Excel dates so they sort correctly. Currency symbols are preserved as cell formatting, not as part of the cell value.

The reverse direction (combining Excel data back into a PDF) is Excel to PDF.

How it works

  1. 1
    Upload your PDF
    Drop your PDF into the upload box. Born-digital PDFs (exported from accounting software, etc.) extract cleanest; scanned PDFs work via OCR but may need touch-ups.
  2. 2
    Run the conversion
    Press Convert. Born-digital PDFs finish in well under a minute; scanned PDFs take 2–4 seconds per page because each page is OCR’d.
  3. 3
    Open in Excel
    You’ll get an .xlsx with each detected table as its own worksheet. Open in Excel, Google Sheets, Numbers, or LibreOffice.
  4. 4
    Touch up if needed
    For scanned sources or complex layouts, scan the output for any rows where columns shifted; manual fix in Excel takes seconds per row.

Real-world uses

Accountants

Bank statements arrive as PDF; extract into Excel for reconciliation in seconds rather than retyping every line.

Financial analysts

Earnings reports published as PDF extract into spreadsheets for modelling without manual data entry.

Researchers

Datasets published in research papers (as PDF tables) extract into Excel for re-analysis.

Bookkeepers

Vendor invoices in PDF format extract into Excel for monthly expense aggregation.

Common questions

How accurate is the extraction?

For born-digital PDFs with clean tables: near-perfect, including correct numeric typing. For scanned PDFs: typically 95–99% on modern, clean scans — the bottleneck is OCR accuracy on the underlying image. Expect to spend a minute reviewing the output spreadsheet for any rows where columns shifted, especially on heavily-formatted source tables.

Do numbers come through as numbers, not text?

Yes. Dollar amounts, percentages, integers, and dates are detected and typed appropriately in the output spreadsheet — you can immediately SUM() columns, sort by date, and use formulas. Currency symbols carry over as cell formatting, not as part of the cell value.

What happens to multi-table PDFs?

Each detected table becomes its own worksheet in the output .xlsx, named by position in the source document. A bank statement with summary + transaction sections produces a workbook with at least those two sheets.

Does it work on scanned PDFs?

Yes. Scanned PDFs are routed through OCR first, then table detection. Accuracy depends on scan quality — phone photos in good light produce near-as-clean output as digital sources; heavily skewed or low-contrast scans need more touch-up.

Can I extract one specific table instead of all of them?

Not in one step — PDF to Excel extracts all detected tables. Workaround: use Extract Pages to pull only the pages containing the table you want, then run that smaller PDF through PDF to Excel.

Is there a page limit?

No hard cap — only the upload-size cap (25 MB free, 1 GB Pro). Long documents take longer to process: budget 1–2 seconds per page for born-digital, 2–4 for scanned.

Related tools & guides