Urgent.News

What's breaking now, across thousands of outlets.

Tech

OCR It – pull text out of un-copyable documents for your LLM

A Chrome extension simplifies extracting text from paginated documents trapped in viewers, such as scanned books, slide decks, and PDFs. Once installed, drag a capture region, assign a hotkey, and hit it on every page to screenshot the exact area, OCR it, and append the text to a running transcript. Alternatively, let the extension handle the entire process with a single hotkey: capture, turn the page, repeat until the document ends.

The extension runs OCR locally using a bundled Tesseract build, requiring no API keys or network connections, ensuring data privacy and security. After each capture, the extension lists the pages with thumbnails, allowing users to identify any drift in the region. Text is editable in place, and bad reads can be re-run individually.

The extension supports auto-page turning after capture and provides progress reports, ensuring reliable end-detection for unattended loops. Supported languages include English, Portuguese, Spanish, and other ~100 languages via vendor-provided models.

Written by urgent.news from Hacker News's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at github.com →

More in Tech

More from Monday 24 August →