AI Tools / AI products & services

OCR It

OCR It is a Chrome and Firefox extension for extracting text from paginated document viewers, including scanned books, slide decks, and PDFs where text selection is unavailable. Users pin a screen region, capture it with a hotkey, and append the OCR result to an ordered transcript; an automatic mode captures each page, advances the viewer, and stops when text repeats, paging fails, OCR fails, or a page limit is reached.

OCR runs locally through a bundled Tesseract build. The extension crops and optionally enlarges and converts the visible region to grayscale, queues OCR jobs serially, stores page text and verification thumbnails, flags consecutive duplicate pages, and exports the transcript as copied text or a TXT file with page separators. Page turning can use a picked screen point or a keyboard event, including controls inside supported frames and shadow roots.

English, Portuguese, and Spanish models are included, with support for adding other Tesseract languages at build time. It makes no outbound requests and requires no API key; local files and some cross-origin viewers require additional browser permissions, and Chrome's built-in PDF viewer supports manual paging rather than automatic advancement. The source is released under the MIT license.

Other tool — not AI
View repository
Mentioned in 1 video
Kind Other tool — not classified as AI

What people use it for

Where it was mentioned