Open the Vootkit workspace.
01PDF & Image OCR
Extract text from scanned PDFs and images with OCR.
- 1Choose toolOpen the Vootkit workspace.
- 2Add inputProvide the content or settings.
- 3ProcessLet the browser do the work.
How to use PDF & Image OCR
Provide the content or settings.
02Let the browser do the work.
03Copy, download or continue.
04Your files stay private
Your work is processed locally in your browser where possible and is never added to a Vootkit upload library.
Learn more about privacyA scanned page looks like text and behaves like a photograph. You cannot search it, select from it, or paste a line into an email — because as far as the file is concerned there are no words on it, only pixels arranged suggestively.
What PDF & Image OCR does
Runs Tesseract optical character recognition in your browser, rendering each PDF page at 2× scale first because recognition accuracy depends heavily on the resolution it is given.
Six languages are available, and choosing the right one matters more than people expect — an English model reading French will mangle every accent.
Engine and languages
| Engine | Tesseract, running locally in your browser |
|---|---|
| Languages | English, Spanish, French, German, Italian, Portuguese |
| Render scale | 2× the PDF page size, for accuracy |
| Accepts | PDF or image files |
| Output | Plain text you can copy |
| First run | Downloads the OCR engine, then works from cache |
| Privacy | The file never leaves your device |
| Handwriting | Not supported — printed text only |
Detailed steps
- Add the scanned PDF or photograph.
- Choose the language of the document, not your own.
- Run it — the first run downloads the engine, which takes a moment.
- Copy the text out and proofread it. OCR is never perfect.
Worth knowing
Scan quality decides everything and no setting compensates for a bad source. A straight, well-lit 300 DPI scan reads almost perfectly; a phone photo taken at an angle in poor light will produce nonsense whatever you do. If the result is bad, rescan rather than retry — and straighten the page first.
Frequently Asked Questions
Why is the text full of mistakes?
Almost always the source. OCR needs sharp, straight, well-lit text; skew, shadow, low resolution and JPEG artefacts each cost accuracy, and they compound. Unusual fonts and tables also read poorly because the layout confuses line detection.
Can it read handwriting?
No. Tesseract recognises printed characters, and handwriting recognition is a genuinely different problem. Even neat printing by hand will produce poor results.
Is my document uploaded to an OCR service?
No — this is unusual and worth stating. The engine is downloaded to your browser and runs there, so the scan never leaves your device. That matters for the things people most often scan: contracts, medical letters, ID documents.
Why is the first run slow?
It downloads the recognition engine and language data once. After that it is cached, and later runs start immediately.
Is PDF & Image OCR free?
Yes. The Vootkit free plan includes 5 tool runs a day. Upgrade to Vootkit Pro for unlimited daily use, an ad-free workspace and saved workflows.
Are my files uploaded?
No. PDF & Image OCR runs entirely in your browser — your file is processed on your own device and never sent to a server. There is nothing for us to store or delete.
Do I need to install anything?
No. PDF & Image OCR works in any modern browser on desktop, tablet or phone. Open the page and start.
How often can I use it? Is there a daily limit?
On the free plan you get 5 tool runs a day. When you reach the limit you'll see a prompt to upgrade, and it resets the next day. Vootkit Pro removes the cap entirely for unlimited daily use.
Recently viewed
This tool processes everything locally in your browser. You can disconnect from the internet after the page loads and it will still work.