Skip to content

9 tools

PDF OCR

Read scanned pages without sending them to a server.

OCR turns a picture of words into actual characters. It is the one genuinely heavy operation on this site: a recognition engine compiled to WebAssembly plus a trained language model, several megabytes in total, downloaded on demand and then cached. That cost buys something unusual — scanned documents, which are often the most sensitive files people have, can be read without ever being uploaded anywhere.

About PDF OCR

Which languages are supported?

English, French, Spanish, German, Italian, Portuguese and Arabic, using the corresponding trained Tesseract models. Each is downloaded only when you select it.

Why is OCR slower than the other tools?

Recognition is computationally expensive and runs on your own CPU rather than a server farm. Expect a few seconds per page, and a one-time model download the first time you use a language.