PDF to text
Extract supported PDF text while keeping lines and columns apart. Uncertain characters are marked instead of guessed.
Drop your file
or drag them here
How it works
- Lines are rebuilt from each word's position, so tables and columns stay apart.
- Supported source text is decoded from the PDF. Incomplete or unreliable mappings are shown with warnings or unknown characters, instead of being silently guessed. See the language matrix for reading limits.
- Detects scanned pages and tells you, instead of returning an empty file.
- Optional JSON output includes every text span with page coordinates.
- Arabic and Hebrew text can be read, but right-to-left reading order is not reconstructed yet.
Need to change text, add a signature or redact something in a PDF?
Open Edit PDFQuestions
Are my files uploaded?
No. This tool runs inside your browser on this device. ORDO's servers never receive the file, and the page's security policy blocks it from sending file data anywhere.
Why is my PDF reported as having no text?
It's a scan or a photo saved as PDF: the pages are images, so there's no text to extract. Reading text from scans requires OCR.