OCR text extraction from images and PDFs

A free browser-based OCR tool for extracting text from images and PDFs. Runs the full Tesseract engine locally via WebAssembly, supports more than 100 languages, fixes sideways photos and shadowed scans automatically, and exports to TXT, Word, Excel or a searchable PDF with an invisible text layer.

How to extract text from an image or PDF

1
Add files

Drop images or PDFs, shoot with the camera on mobile, paste from clipboard with Ctrl+V, or paste an image URL.

2
Pick the language

The default covers the site language plus English; choose up to three languages from the grouped dropdown for other documents.

3
Choose output style

Plain text for copy-paste, Preserve layout for documents, Extract tables for spreadsheets, or Receipt for totals, dates and vendors.

4
Review and edit

Inspect the bounding boxes, fix any low-confidence words inline, search for specific words.

5
Export or copy

Download as TXT, DOCX, XLSX or searchable PDF — or bundle a whole batch into a ZIP.

Turn scans, photos and PDFs into editable text, Word, Excel or searchable PDF

Drop images or PDFs here or browse

Or press Ctrl+V to paste a screenshot · type / paste URL

Working…

High confidence Medium Low — review

Click a box to jump to the word · click the word to highlight the box · Edit any word inline — your corrections are saved to the export.

QR code detected:

Export

Runs in your browser via WebAssembly. Works offline after the first use.

Keyboard shortcuts

Ctrl+VPaste an image from the clipboard Ctrl+FFind in extracted text Ctrl+EnterStart extraction EscCancel current operation [ / ]Jump to next low-confidence word Ctrl+scrollZoom the image preview
Frame the document inside the guide · hold steady · tap to capture
Published Updated Authors: