Description
OCR - Image Reader is a powerful optical character recognition (OCR) extension for Chrome designed to capture and convert images into editable text. Upon installation, it adds a toolbar button to your browser. When activated, users can select a specific region within the active window for OCR processing.
The extension utilizes the Tesseract OCR engine, integrated via the "tesseract.js" library, which boasts support for over 100 languages. Key features include automatic text orientation detection and script recognition, enhancing the accuracy and usability across diverse text sources. The library is loaded on demand and removed after use, ensuring no long-term resource consumption.
Two OCR engines are available: Tesseract, which extracts plain text without preserving formatting, and IBM Granite-Docling, capable of interpreting headings, lists, styles (bold, italics), tables, and mathematical equations. The extension provides a progress bar during detection due to the inherent slowness of OCR processes. Importantly, all OCR processing is performed offline, with no server-side interaction, except for the initial download of language training data.
This tool is versatile, allowing users to extract text from images, PDF documents, PowerPoint slides, or even web pages where content selection might be restricted. For improved accuracy, the extension includes a feature to invert images and retry OCR, which is particularly beneficial for dark-themed interfaces. Users can also manually modify images and re-upload them for another attempt if initial extraction confidence is low.
Version 0.2.4 introduced support for browser-level page zooming and OS-level screen zooming, along with beta-level image language detection. The extension is offered by brian.girko and has garnered a 4.1 out of 5 rating from 243 users, indicating a generally positive reception for its OCR capabilities.
OCR's Core Features
Optical Character Recognition (OCR) for image-to-text conversion
Internal OCR engine (Tesseract via tesseract.js)
Supports over 100 languages
Automatic text orientation and script detection
Offline OCR processing
Choice between Tesseract (plain text) and IBM Granite-Docling (formatted text) engines
Image inversion and retry option for improved accuracy
Ability to extract text from images, PDFs, and Powerpoint slides
Extracts text from web pages with restricted content selection
On-demand loading and unloading of JS library
Progress bar for detection modules
Support for browser and OS-level zooming
Beta image language detection
How to use OCR?
Install the OCR - Image Reader extension from the Chrome Web Store.
Click the extension's toolbar button to activate it.
Select the desired region on your screen containing the image or document.
The extension will capture the area and perform OCR processing.
Review the extracted text for accuracy.
If needed, modify the image and re-upload it for another OCR attempt.
OCR's Use Cases
- Extract text from images
- Digitize documents
- Process PDF content
- Capture web page text
- Analyze presentation slides
- Improve accessibility
- Data entry automation







