awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 dépôts

Awesome GitHub RepositoriesMultimodal Information Extractors

Engines designed to parse and structure information from mixed-media document formats.

Distinguishing note: Focuses on the parsing engine aspect of multimodal extraction, distinct from the broader extraction frameworks.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Multimodal Information Extractors. Refine with filters or upvote what's useful.

Awesome Multimodal Information Extractors GitHub Repositories

Trouvez les meilleurs dépôts grâce à l'IA.Nous recherchons les dépôts les plus pertinents grâce à l'IA.
  • hkuds/lightragAvatar de HKUDS

    HKUDS/LightRAG

    36,651Voir sur GitHub↗

    LightRAG is a graph-based retrieval framework designed to build retrieval-augmented generation pipelines. It structures unstructured text into knowledge graphs, enabling multi-hop reasoning and complex query synthesis across large document collections. By integrating dense vector embeddings with structured knowledge graphs, the system facilitates both similarity-based and relationship-aware information retrieval. The framework distinguishes itself through a dual-level retrieval strategy that combines low-level keyword matching with high-level semantic graph traversal to capture both specific

    A processing engine that parses both text and visual data from diverse document formats to build comprehensive searchable knowledge bases.

    Pythongenaigptgpt-4
    Voir sur GitHub↗36,651
  • codexu/note-genAvatar de codexu

    codexu/note-gen

    12,173Voir sur GitHub↗

    Note-gen is an artificial intelligence-assisted note-taking application and knowledge management tool designed for local-first data ownership. It functions as a workspace that leverages language models to organize, summarize, and synthesize personal notes into structured documents while maintaining offline accessibility. The platform distinguishes itself through a multimodal workflow orchestrator that chains sequences of tasks to process text, images, and external data. By integrating vision-language models, it extracts information from visual inputs like screenshots and documents, converting

    Parses and structures information from screenshots and documents using vision-language models.

    TypeScriptagentchatbotknowledge-base
    Voir sur GitHub↗12,173
  1. Home
  2. Artificial Intelligence & ML
  3. Multimodal Information Extractors