awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com

OCR screen capture

Ranking updated Sep 7, 2026

For screen text extractors, the first results are xushengfeng/esearch (This desktop tool provides local and cloud-based optical character recognition alongside screen capture, region selection, and clipboard integration, making it a comprehensive solution for extracting text from your display), pot-app/pot-desktop (This cross-platform desktop utility provides screen capture extraction, optical character recognition, global hotkeys, and clipboard integration as a background service, though its primary focus leans towards translation workflows) and pantsudango/dango-translator (Dango-Translator captures text from screens using optical character recognition, though its primary focus is on immediate translation rather than general text copying). sharex/sharex and microsoft/powertoys round out the shortlist. Compare the match explanations and check the project documentation against your requirements.

Hand-picked open-source screen text extractors and OCR tools ranked by stars and activity. Compare the top alternatives and pick the right one.

OCR screen capture

Find the best repos with AI.We'll search the best matching repositories with AI.
  • xushengfeng/esearchxushengfeng avatar

    xushengfeng/eSearch

    6,275View on GitHub↗

    eSearch is a desktop tool that combines screen capture, image annotation, screen recording, optical character recognition (OCR), and text search and translation into a single application. It is built around a modular architecture that coordinates these tasks through an event-driven capture pipeline, allowing users to capture screen regions, annotate them with drawing and shape tools, and then extract text using a local-first OCR engine or optional cloud services. The project distinguishes itself by integrating a command-line interface for triggering capture and recognition tasks, enabling scr

    This desktop tool provides local and cloud-based optical character recognition alongside screen capture, region selection, and clipboard integration, making it a comprehensive solution for extracting text from your display.

    TypeScriptOptical Character RecognitionScreen Capture Tools
    View on GitHub↗6,275
  • pot-app/pot-desktoppot-app avatar

    pot-app/pot-desktop

    17,110View on GitHub↗

    This application is a cross-platform desktop utility designed for automated translation, optical character recognition, and speech synthesis. It functions as a modular client that integrates various local and remote language services, allowing users to process text through hotkeys, clipboard monitoring, or direct input. The software distinguishes itself through a plugin-based architecture and a built-in automation framework. By exposing a local network interface, it enables external applications and scripts to programmatically trigger its translation and recognition workflows. Users can furth

    This cross-platform desktop utility provides screen capture extraction, optical character recognition, global hotkeys, and clipboard integration as a background service, though its primary focus leans towards translation workflows.

    JavaScriptOptical Character RecognitionGlobal Hotkey ManagersSystem Clipboard Access
    View on GitHub↗17,110
  • pantsudango/dango-translatorPantsuDango avatar

    PantsuDango/Dango-Translator

    8,411View on GitHub↗

    Dango-Translator is an OCR translation system and multi-engine translation client designed to extract text from images or screens and replace it with translated content. It functions as an image text translator and real-time screen translator, utilizing optical character recognition to convert text between different languages automatically. The software distinguishes itself through coordinate-based image typesetting and a glossary manager. These tools allow for the replacement of original image content with translated text in the same area and the use of specialized dictionaries to ensure con

    Dango-Translator captures text from screens using optical character recognition, though its primary focus is on immediate translation rather than general text copying.

    PythonOptical Character Recognition
    View on GitHub↗8,411
  • sharex/sharexShareX avatar

    ShareX/ShareX

    38,123View on GitHub↗

    ShareX is a desktop utility designed for screen capture, image annotation, and automated file sharing. It provides a comprehensive suite of tools for capturing screen regions, windows, or scrolling content, and includes a layered image editor that allows users to manipulate, scale, and transform graphical elements and annotations directly on captured media. The application distinguishes itself through an event-driven post-capture pipeline that triggers automated workflows, such as image processing, external command execution, or file uploads, immediately after a capture event. Users can exten

    ShareX is a feature-rich screen capture and workflow utility that includes built-in optical character recognition and clipboard integration, making it a capable tool for extracting screen text despite its primary focus on screenshots and file sharing.

    C#Optical Character RecognitionScreen Capture Tools
    View on GitHub↗38,123
  • microsoft/powertoysmicrosoft avatar

    microsoft/PowerToys

    135,047View on GitHub↗

    PowerToys is a collection of background-resident system utilities designed to extend native operating system functionality and streamline desktop workflows. It operates as a modular toolkit, utilizing a central plugin-based host architecture that allows users to dynamically enable or disable specific features for system configuration and automation. By leveraging native system hooking, the suite intercepts global input and window events to provide advanced control over the computing environment. The project distinguishes itself through its focus on cross-device input orchestration and spatial

    Microsoft PowerToys includes a built-in screen text extractor utility that performs optical character recognition with hotkey and region selection support, though the repository is a broader system utility suite rather than a dedicated OCR tool.

    CCross-Device Input ControllersDesktop Workflow OptimizersPlugin-Based Architectures
    View on GitHub↗135,047
  • ripperhe/bobripperhe avatar

    ripperhe/Bob

    9,693View on GitHub↗

    Bob is an extensible macOS utility designed for screen text extraction, translation aggregation, and speech synthesis. It functions as a wrapper that integrates multiple optical character recognition and translation services into a single interface, allowing users to capture screen areas, decode QR codes, and convert visual text into editable strings. The tool distinguishes itself through a plugin-based architecture that supports the integration of custom translation, speech synthesis, and image recognition APIs. It enables multi-engine parallel execution, allowing a single request to be proc

    Bob is an extensible screen text extraction utility that integrates optical character recognition for capturing and decoding visual text, though it is specifically tailored for macOS rather than being fully cross-platform.

    Screen Capture and Text ExtractionText Translation ServicesAutomated Translation Workflows
    View on GitHub↗9,693
  • miaomiaosoft/pandaocrmiaomiaosoft avatar

    miaomiaosoft/PandaOCR

    5,274View on GitHub↗

    PandaOCR is a desktop application for extracting text from images and screen captures using optical character recognition. It functions as a mathematical formula digitizer, a table data extractor, a multilingual translation utility, and a text-to-speech interface. The project distinguishes itself through specialized recognition routing that distributes data across different providers based on whether the content is standard text, tables, or formulas. It provides real-time software interface localization by rendering translated text layers directly over active application windows using coordin

    PandaOCR is a desktop application for capturing screen text and extracting it via optical character recognition, making it a fitting tool for this search despite lacking explicit mentions of global hotkeys or cross-platform support.

    Image Text ExtractionsText RecognitionAutomated Translators
    View on GitHub↗5,274
  • copytranslator/copytranslatorCopyTranslator avatar

    CopyTranslator/CopyTranslator

    17,749View on GitHub↗

    CopyTranslator is a clipboard-based translation tool and multi-engine translation client that monitors the system clipboard to provide automatic language conversion. It functions as an assistant that integrates large language models, cloud translation APIs, and digital dictionaries to produce context-aware translations and side-by-side reading views. The application includes a specialized PDF text cleaner to remove formatting artifacts and line breaks from copied content. It also features an optical character recognition extractor to convert images or screen captures into editable text for im

    CopyTranslator includes an optical character recognition extractor for screen captures, though its primary focus is clipboard translation rather than general text extraction.

    TypeScriptAI-Triggered Clipboard ActionsClipboard TranslatorsAI Translation Assistants
    View on GitHub↗17,749
  • thejoefin/text-grabTheJoeFin avatar

    TheJoeFin/Text-Grab

    4,610View on GitHub↗

    Text-Grab is a desktop utility that captures text from screen regions, images, PDFs, and native user interface elements using on-device optical character recognition (OCR) and Windows UI Automation. It processes text entirely locally without sending data to external services, and extracts text directly from UI controls with perfect accuracy by reading the accessibility tree. The application also includes a persistent snippet dictionary for instant retrieval of frequently used text via a configurable system-wide hotkey. The tool supports building reusable extraction workflows by saving capture

    Text-Grab is a local screen-capture utility with optical character recognition and global hotkey support, though it is limited to Windows rather than being cross-platform.

    C#OCRBatch Folder ProcessorsLocal-First Engines
    View on GitHub↗4,610
  • soffes/hotkeysoffes avatar

    soffes/HotKey

    1,077View on GitHub↗

    HotKey is a developer library for registering system-wide keyboard shortcuts in macOS applications. It binds modifier keys and key codes to system-wide identifiers using native operating system event hooks, allowing applications to respond to key combinations while running in the background. The library executes custom closures asynchronously when native keyboard monitor notifications match registered global shortcut signatures. It handles both press and release events, and automatically unregisters system hot key bindings during object deallocation to prevent stale event listeners and memory

    This project is a developer library for registering system-wide shortcuts on macOS rather than a complete screen text extractor tool, serving as a low-level building block rather than the ready-to-use application requested.

    SwiftGlobal Keyboard ShortcutsGlobal Hotkey ManagersGlobal Keyboard Shortcuts
    View on GitHub↗1,077
  • zenorocha/clipboard.jszenorocha avatar

    zenorocha/clipboard.js

    34,140View on GitHub↗

    clipboard.js is a lightweight JavaScript library and browser API wrapper designed to manage text transfers to the system clipboard. It functions as a DOM event clipboard manager that enables the copying and cutting of text from web page elements. The library provides mechanisms for dynamic text transfer, allowing text to be resolved from static HTML elements, specific data attributes, or programmatically defined strings at runtime. It includes an event-driven callback system to trigger user interface feedback or custom actions upon the success or failure of a clipboard operation. The tool im

    This repository provides a JavaScript library for managing browser clipboard actions rather than an OCR tool for capturing text from a computer screen, making it a different kind of utility altogether.

    JavaScriptClipboard CopyingClipboard Management
    View on GitHub↗34,140
Compare the top 10 at a glance
RepositoryStarsLanguageLicenseLast push
xushengfeng/esearch6.3KTypeScriptgpl-3.0Feb 19, 2026
pot-app/pot-desktop17.1KJavaScriptgpl-3.0Jan 21, 2026
pantsudango/dango-translator
8.4K
Python
lgpl-2.1
Feb 15, 2026
sharex/sharex38.1KC#GPL-3.0Jun 16, 2026
microsoft/powertoys135KCMITJun 16, 2026
ripperhe/bob9.7K——Dec 30, 2025
miaomiaosoft/pandaocr5.3K——Aug 28, 2022
copytranslator/copytranslator17.7KTypeScriptGPL-2.0Feb 23, 2026
thejoefin/text-grab4.6KC#mitFeb 12, 2026
soffes/hotkey1.1KSwiftMITDec 29, 2024

Related searches

  • Text conversion tools
  • Text search engine
  • Text metric calculator
  • an ocr tool for extracting text
  • Table extraction tools
  • Text editors
  • a command line tool for text processing
  • an OCR engine for document text extraction