awesome-repositories.com
Blog
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectDespreCum realizăm clasamentulPresăServer MCP
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
miaomiaosoft avatar

miaomiaosoft/PandaOCR

0
View on GitHub↗
5,274 stele·660 fork-uri·3 vizualizări

PandaOCR

PandaOCR este o aplicație desktop pentru extragerea textului din imagini și capturi de ecran folosind recunoașterea optică a caracterelor (OCR). Funcționează ca un digitizor de formule matematice, extractor de date din tabele, utilitar de traducere multilingvă și interfață text-to-speech.

Proiectul se remarcă printr-un sistem de rutare specializat care distribuie datele către diferiți furnizori, în funcție de tipul conținutului: text standard, tabele sau formule. Oferă localizarea interfeței software în timp real, randând straturi de text tradus direct peste ferestrele active ale aplicațiilor, folosind elemente flotante aliniate prin coordonate.

Capabilitățile extinse includ recunoașterea textului în loturi (batch), cu detectare automată a limbii și recompunere euristică a textului pentru a îmbina fragmentele în propoziții coerente. Instrumentul suportă, de asemenea, automatizarea prin monitorizarea clipboard-ului și posibilitatea de a salva coordonate fixe ale ecranului pentru extragerea repetată din regiuni specifice.

Sistemul se integrează cu servicii externe de traducere și diverse motoare de sinteză vocală pentru a converti caracterele digitale recunoscute în output audio.

Features

  • Image Text Extractions - Extracts editable text from screenshots and images using a variety of optical character recognition engines.
  • Text Recognition - Extracts editable text from screenshots and images using various optical character recognition engines.
  • OCR Engine Routing - Routes image data to specialized providers depending on whether the content is standard text, tables, or formulas.
  • Screen Text Extractors - Desktop application that performs OCR on arbitrary screen regions to capture non-selectable text.
  • Translation API Integrations - Integrates with external translation services via APIs to provide multilingual text conversion.
  • Real-time Software Localization - Translates foreign language software interfaces in real time to enable navigation of non-native applications.
  • Automated Translators - Captures screen text and instantly translates it to help users understand foreign media or documents.
  • Image-Based Table Extractors - Recognizes structured data from images of tables and transforms it into formatted digital text.
  • Structured Spreadsheet Tables - Identifies and extracts structured table data from images and converts it into digital text.
  • Visual Table Extraction - Extracts structured data from images of tables and transforms it into a digital text format.
  • OCR Integration APIs - Uses REST interfaces and wrapper APIs to connect to external optical character recognition services.
  • Translation Utilities - Provides a desktop utility that integrates OCR and translation services for multilingual support.
  • Formula Recognition Engines - Converts images of complex mathematical expressions into editable digital formats using specialized recognition engines.
  • Mathematical Digitization Engines - Converts images of complex mathematical expressions into editable digital formats using specialized engines.
  • Software Interface Localization - Renders translated text layers directly over active application windows to provide real-time software interface localization.
  • Window-Based Overlay Rendering - Renders translated text layers directly over active application windows using coordinate-aligned overlays.
  • Multi-Language Recognition Models - Supports target language selection and automatic language detection to improve text extraction accuracy.
  • Text-to-Coordinate Mapping - Saves specific X and Y pixel offsets to automate text extraction from fixed screen locations.
  • Batch Processing - Enables sequential processing of multiple images for consistent, high-volume text extraction.
  • OCR Integration Gateways - Connects to external OCR providers and registered API interfaces to convert image content into digital text.
  • Clipboard Translators - Monitors the system clipboard to automatically trigger recognition and translation workflows.
  • Screen Capture Extraction - Captures specific screen regions and converts visual content to text for repeated extraction.
  • AI Text Fidelity Refiners - Merges fragmented OCR results into coherent sentences by analyzing spatial layout and reading order.
  • OCR Layout Recomposition - Heuristically merges fragmented OCR results into coherent sentences by analyzing spatial layout.
  • Clipboard Monitoring - Automatically triggers recognition and translation by observing changes to the system clipboard.

Istoric stele

Graficul istoricului de stele pentru miaomiaosoft/pandaocrGraficul istoricului de stele pentru miaomiaosoft/pandaocr

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Întrebări frecvente

Ce face miaomiaosoft/pandaocr?

PandaOCR este o aplicație desktop pentru extragerea textului din imagini și capturi de ecran folosind recunoașterea optică a caracterelor (OCR). Funcționează ca un digitizor de formule matematice, extractor de date din tabele, utilitar de traducere multilingvă și interfață text-to-speech.

Care sunt principalele funcționalități ale miaomiaosoft/pandaocr?

Principalele funcționalități ale miaomiaosoft/pandaocr sunt: Image Text Extractions, Text Recognition, OCR Engine Routing, Screen Text Extractors, Translation API Integrations, Real-time Software Localization, Automated Translators, Image-Based Table Extractors.

Care sunt câteva alternative open-source pentru miaomiaosoft/pandaocr?

Alternativele open-source pentru miaomiaosoft/pandaocr includ: hanmin0822/misakatranslator — MisakaTranslator is a real-time game translation tool designed to extract text from games and manga and provide… paddlepaddle/paddlex — PaddleX is a PaddlePaddle-based framework for building, deploying, and fine-tuning AI model pipelines, with pre-built… pot-app/pot-desktop — This application is a cross-platform desktop utility designed for automated translation, optical character… hillya51/lunatranslator — LunaTranslator is a real-time translation tool designed for visual novels and games. It functions as a multi-engine… ripperhe/bob — Bob is an extensible macOS utility designed for screen text extraction, translation aggregation, and speech synthesis.… optikey/optikey — OptiKey is an assistive technology suite and gaze-based input system designed to provide computer access and…

Alternative open-source pentru PandaOCR

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu PandaOCR.
  • hanmin0822/misakatranslatorAvatar hanmin0822

    hanmin0822/MisakaTranslator

    5,712Vezi pe GitHub↗

    MisakaTranslator is a real-time game translation tool designed to extract text from games and manga and provide machine translations via external engines. It functions as a text extractor using both memory hooking to retrieve raw text directly from running processes and optical character recognition to convert images of in-game text into editable strings. The tool includes a speech synthesizer to read translated dialogue and sentences aloud. To maintain accuracy, it utilizes a custom translation dictionary to manage specialized word lists and manual phrase mappings for character names and loc

    C#comiccsharpgalgame
    Vezi pe GitHub↗5,712
  • paddlepaddle/paddlexAvatar PaddlePaddle

    PaddlePaddle/PaddleX

    6,163Vezi pe GitHub↗

    PaddleX is a PaddlePaddle-based framework for building, deploying, and fine-tuning AI model pipelines, with pre-built support for computer vision, OCR, document analysis, and time series tasks. It offers a toolkit of ready-to-use pipelines for image classification, object detection, segmentation, and pose estimation, alongside an end-to-end OCR document analysis pipeline that extracts text, tables, formulas, and layout information. The platform also includes a dedicated time series forecasting pipeline for analyzing historical data to detect anomalies, classify patterns, and predict future val

    Pythonai-pipelinesclassificationdeployment
    Vezi pe GitHub↗6,163
  • pot-app/pot-desktopAvatar pot-app

    pot-app/pot-desktop

    17,110Vezi pe GitHub↗

    This application is a cross-platform desktop utility designed for automated translation, optical character recognition, and speech synthesis. It functions as a modular client that integrates various local and remote language services, allowing users to process text through hotkeys, clipboard monitoring, or direct input. The software distinguishes itself through a plugin-based architecture and a built-in automation framework. By exposing a local network interface, it enables external applications and scripts to programmatically trigger its translation and recognition workflows. Users can furth

    JavaScriptlinuxmacosocr
    Vezi pe GitHub↗17,110
  • hillya51/lunatranslatorAvatar HIllya51

    HIllya51/LunaTranslator

    12,030Vezi pe GitHub↗

    LunaTranslator is a real-time translation tool designed for visual novels and games. It functions as a multi-engine translation hub and text extractor that captures dialogue via memory hooking or optical character recognition to convert it into a target language. The project distinguishes itself through specialized linguistic tools, including a Japanese text analyzer for sentence segmentation and phonetic readings. It also operates as a digital dictionary aggregator, querying multiple online and offline databases simultaneously to provide comprehensive vocabulary definitions for language lear

    C++galgameocrreverse-engineering
    Vezi pe GitHub↗12,030
Vezi toate cele 30 alternative pentru PandaOCR→