PDF to Text, HTML & XML
Works on scanned PDFs too, with built-in OCR.
How to use
- Choose your PDF.
- Pick TXT, HTML or XML, and the language if it is a scan.
- Press Extract, then Download.
Good to know
PDFs made on a computer are read instantly. Scanned pages are read with OCR, which is slower. The first time, your browser downloads the language data, which can take a minute. OCR works best on clear printed text. It can make mistakes, and it cannot read handwriting well. Always check the result. The output keeps the text, not the original layout or tables.
FAQ
Is my PDF uploaded?
No. Your pages are read inside your browser. Only the OCR language data is downloaded to your device.
Why is OCR slow?
Reading letters from pictures takes a lot of work. A page can take 5 to 30 seconds on a phone. OCR is limited to 20 scanned pages for now.
Can it make a Word file?
No. Converting PDF to Word accurately is not possible in a browser, so we do not offer it.