Read positioned text
The utility reads extractable text and its page position. It does not send the PDF to a conversion server.
Free · Private · No account
Status: Live Version: 1.0
Turn tables in text-based PDFs into spreadsheet-ready CSV or tab-delimited files without uploading the document.
Step 1
One PDF, up to 50 MB and 300 pages. Scans, OCR, forms, and password-protected files are not supported.
Your file is read only after you choose it and is processed locally.
Step 2 · Review
Structured layout recognized
A high-confidence document structure was recognized. This optional export combines related sections, normalizes consistent fields, and excludes repeated headers and non-data rows. Review the result against the source document.
Raw table export
Exports include the full cleaned table—not only the rows currently visible.
How it works
The utility reads extractable text and its page position. It does not send the PDF to a conversion server.
Rows and columns are inferred from recurring spacing, headers, and data patterns across pages.
Detection is not a guarantee. Check dates, signs, balances, wrapped descriptions, and continuation pages before using the export.
Scanned PDFs need OCR and are not supported. The utility does not edit individual cells, make open-ended content decisions, create XLSX files, store document history, or provide professional advice. Smart cleaning appears only when a supported structured layout is recognized with high confidence.
Using this utility
This utility is for people who need spreadsheet-ready rows from a structured PDF, invoice, report, price list, inventory sheet, schedule, or government table without uploading the document.
Supporting guidance updated .
The extractor infers rows and columns from positioned text. Cleanup choices change the exported view, while dates, currency symbols, signs, balances, and source order remain text exactly as detected.
Validation checklist:
A clean-looking preview is evidence to inspect, not proof of accuracy. If the document is a scan or the inferred structure remains ambiguous, stop and use OCR or manual review suited to the stakes.
PDFs do not necessarily store semantic tables, so a plausible preview can still be wrong. Scans require OCR and are not supported. Always compare row counts, signs, wrapped descriptions, and continuation pages with the source.
No. The utility parses it in a browser worker and does not send document bytes or extracted values to Nichessities.
A scan stores page images. OCR is intentionally outside this version.
The apostrophe prevents untrusted PDF text from being interpreted as a spreadsheet formula.
No. Cleanup is limited to headers, empty rows or columns, repeated headers, row deletion, and likely wrapped text. Use a spreadsheet for cell editing.