0 1 0 1 0 1 0 1 0 1 0 1 0 1 0 1 0 1 0 1 0 1 0 1 0
OCR: image and PDF to text
Extract the text from an image, screenshot or PDF. For PDFs it combines the document's own text —including signatures and fields— with optical character recognition (OCR) of the images. Everything runs in your browser, without uploading the file to any server.

Drag an image or a PDF, or click to select

PDF, PNG, JPEG, WebP, GIF, BMP

Choose one or more languages present in the image

How it works

OCR (optical character recognition) turns the text that appears inside an image or a PDF —a photo, a screenshot or a scanned document— into editable text that you can copy, search or paste anywhere.

You can choose one or more languages to improve accuracy: setting the correct language helps the engine better interpret accents, special characters and whole words. The first time you use a language its model is downloaded and then stored for next time.

All recognition happens in your browser with WebAssembly: the file is not uploaded to any server. The result includes a confidence estimate so you know how reliable the reading was.

Use cases

  • Copy the text from a screenshot when you can't select it.
  • Digitize a scanned document, invoice or receipt to edit it.
  • Extract a code snippet or an error message from a photo.
  • Recover text from an image to translate or search it.
  • Extract the text from a PDF, whether digital or scanned, to reuse it.

Frequently asked questions

Is my image uploaded to any server?

No. Recognition runs entirely in your browser; the image never leaves your device. Only the language models you choose are downloaded, once.

Why should I choose the text language?

Each language has its own trained model. Selecting the right language noticeably improves accuracy, especially with accents and language-specific characters. If an image mixes languages, you can select several at once.

How do I get better results?

Use sharp, well-lit images with good contrast between text and background. Horizontal, undistorted text works better than angled, blurry or handwritten photos.

What formats can I use?

You can use images in the usual formats (PNG, JPEG, WebP, GIF and BMP) as well as PDF files. For documents, a screenshot or scan at good resolution gives better results.

How are PDFs processed?

They are processed page by page, combining all available text: the PDF's own text —including form fields and digital signatures— is extracted directly and exactly, and the text inside the images is also recognized with OCR. "Digital" PDFs are read instantly; scanned ones or those with many images may take longer.