ParseJet

PDF to Excel — Extract Tables from PDF

Upload a PDF and download an Excel workbook with its tables. ParseJet detects the tables on every page, keeps rows and columns intact, turns figures into real numbers, and puts each table on its own sheet. Works on invoices, bank statements, reports and data exports; scanned PDFs need the PDF OCR tool first.

Drop a file here or browse

Accepts PDF files

Free — 3 requests/day, no signup. for 300 credits/month free.

Files are converted on ParseJet's servers and deleted right after the response.

How it works

1

Upload your PDF

Drop a PDF that contains tables — invoices, statements, price lists, reports.

2

Tables are detected

ParseJet finds each table by its column structure and ruling, reads the cells and types the numbers.

3

Download the .xlsx

One sheet per table, named by page. Open it in Excel, Google Sheets or Numbers — or get CSV instead through the API.

Key features

What makes this pdf to excel stand out.

All tables, every page

Multi-page documents are handled in one go; each table gets its own sheet named after its page.

Real numbers, not text

1,250.00 becomes 1250 so totals, sorting and formulas work immediately.

Headers kept

The first row of each table is written in bold and column widths are fitted to the content.

CSV when you want it

The same conversion returns CSV for scripts and imports — ask for output_format=csv in the API.

No watermark, no email

Download the file straight away. Nothing is added to it and nothing is kept on our side.

Batch via API

Convert hundreds of statements with a few lines of Python or JavaScript.

Use cases

Common scenarios where this tool saves you time.

Bank and card statements

Get transactions into a spreadsheet for budgeting, reconciliation or tax time.

Invoices and price lists

Line items and prices land in cells you can sum and compare.

Reports and research

Lift the data tables out of a PDF report to chart or analyse them.

Data pipelines

Extract tables from incoming PDFs automatically and load them into a database.

Automate with the API

Use the same tool programmatically. Works with any language — just HTTP.

cURL
# PDF tables → Excel workbook (one sheet per table)
curl -X POST https://api.parsejet.com/v1/convert/document \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -F "[email protected]" -F "output_format=xlsx" \
  -o statement.xlsx

# ... or CSV
curl -X POST https://api.parsejet.com/v1/convert/document \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -F "[email protected]" -F "output_format=csv" -o statement.csv
Python
import httpx

resp = httpx.post(
    "https://api.parsejet.com/v1/convert/document",
    headers={"Authorization": "Bearer YOUR_API_KEY"},
    files={"files": open("statement.pdf", "rb")},
    data={"output_format": "xlsx"},
    timeout=120,
)
resp.raise_for_status()
open("statement.xlsx", "wb").write(resp.content)
print(resp.headers["x-tables"], "tables")
JavaScript
const formData = new FormData();
formData.append("files", pdfFile);
formData.append("output_format", "xlsx");

const res = await fetch("https://api.parsejet.com/v1/convert/document", {
  method: "POST",
  headers: { Authorization: "Bearer YOUR_API_KEY" },
  body: formData,
});
const blob = await res.blob(); // the .xlsx file
console.log(res.headers.get("x-tables"), "tables");

Want to automate this?

ParseJet API gives you the same parsing power via a single HTTP endpoint. No ffmpeg, no poppler, no tesseract — just one API call.

curl -X POST https://api.parsejet.com/v1/parse/auto/url \ -H "Content-Type: application/json" \ -d '{"url":"https://example.com"}'
Read API Docs

Frequently asked questions

How do I convert a PDF to Excel?

Upload the PDF above and click Convert. You get an .xlsx file with one sheet per table found in the document. Three conversions a day are free without an account.

Does it work with scanned PDFs?

Table detection needs a text layer. If your PDF is a scan or a photo, run it through the PDF OCR tool first; the table structure then has to be visible as ruled lines or aligned columns for detection to pick it up.

Why did it say no tables were found?

The PDF either has no text layer (scan) or its data is laid out as free text rather than a grid. Try PDF OCR for scans; for text-only documents, the PDF to Text tool gives you the content to reshape by hand.

Will numbers be numbers in Excel?

Yes. Cells that look like numbers — including 1,250.00 and -3.5 — are written as numeric values so you can sum and sort. Text, dates and mixed cells stay as text.

Can I get CSV instead?

Yes. Through the API, set output_format=csv; all tables are written one after another with a blank line between them.

Is it free?

Yes. Three free conversions a day with no signup; a conversion costs 2 credits and the free account includes 300 credits per month. Paid plans start at $19/month.

Start extracting text for free

No signup required. Parse your first file in seconds.

View Pricing