awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 dépôts

Awesome GitHub RepositoriesVisual Context Awareness Engines

Systems that extract visual metadata from the desktop to provide situational awareness to language models.

Distinct from Context-Aware Tooling: Distinct from general context-aware tooling: focuses on visual desktop metadata extraction for LLM situational awareness.

Explore 5 awesome GitHub repositories matching software engineering & architecture · Visual Context Awareness Engines. Refine with filters or upvote what's useful.

Awesome Visual Context Awareness Engines GitHub Repositories

Trouvez les meilleurs dépôts grâce à l'IA.Nous recherchons les dépôts les plus pertinents grâce à l'IA.
  • moeru-ai/airiAvatar de moeru-ai

    moeru-ai/airi

    41,040Voir sur GitHub↗

    Airi is an interactive digital companion engine designed to bridge large language models with local animation rendering. It functions as a middleware platform that synchronizes conversational text streams with skeletal and facial movements to drive virtual avatars in real time. The framework distinguishes itself by integrating desktop context awareness, allowing characters to maintain situational awareness of a user's screen activity across both desktop and web environments. It utilizes a hybrid execution model that splits computational workloads between cloud-based language processing and lo

    Extracts visual metadata from the user environment to provide language models with situational awareness of the active desktop.

    TypeScriptai-companionai-vtuberclawdbot
    Voir sur GitHub↗41,040
  • zuodaotech/everyone-can-use-englishAvatar de ZuodaoTech

    ZuodaoTech/everyone-can-use-english

    34,948Voir sur GitHub↗

    This project is an AI-powered English education tool and browser extension designed for immersive language learning. It functions as a proficiency training suite that integrates AI tools and linguistic analysis directly into external websites and streaming video platforms. The system employs a DOM injection model to add interactive overlays and toolbars to web pages. It uses large language model APIs to provide real-time translations and maps curated pronunciation and grammar exercises to specific timestamps in external media. The tool covers a broad range of English proficiency training, in

    Uses page metadata and active media content to provide real-time translations and linguistic analysis via LLMs.

    TypeScript
    Voir sur GitHub↗34,948
  • mediar-ai/screenpipeAvatar de mediar-ai

    mediar-ai/screenpipe

    19,337Voir sur GitHub↗

    Screenpipe is a local screen and audio recorder that captures and indexes digital activity to create a searchable archive of computer usage. It functions as an AI context engine, providing a local database of visual and auditory history to ground large language models. The system serves as a Model Context Protocol server, delivering screen history and meeting transcriptions to external AI assistants. It utilizes an OCR screen search tool to extract text from visual data and a speech-to-text transcription tool for identifying speakers in system and microphone audio. The software includes capa

    Functions as a visual context awareness engine that grounds LLMs with a local database of visual and auditory history.

    Rust
    Voir sur GitHub↗19,337
  • josstorer/chatgptboxAvatar de josStorer

    josStorer/chatGPTBox

    10,738Voir sur GitHub↗

    chatGPTBox is a browser extension that integrates large language model chat interfaces and AI tools directly into the web browsing experience. It functions as an AI productivity toolkit and API client, allowing users to access AI assistants via a floating chat interface without leaving their active webpage. The project distinguishes itself by offering context-aware assistance and website-specific adaptations based on the current URL. It further enhances the browsing experience by displaying AI-generated responses alongside standard search engine results and providing a system to route chat re

    Uses site-specific configuration mappings to trigger tailored AI behaviors based on the current webpage URL.

    JavaScript
    Voir sur GitHub↗10,738
  • sylinko/everywhereAvatar de Sylinko

    Sylinko/Everywhere

    6,093Voir sur GitHub↗

    Everywhere is a desktop AI assistant that understands whatever is on your screen and can act across applications without requiring screenshots or manual context switching. It reads structured UI data through accessibility and automation APIs to perceive the active application and visible content, then provides context-aware help, summaries, translations, and answers to natural language questions about what you are viewing. The tool distinguishes itself by combining on-screen content analysis with a multi-LLM agent platform that routes requests to providers like OpenAI, Anthropic, and local mo

    Triggers pre-configured actions or agent workflows based on the active application and visible content.

    C#aiai-agentsai-assistant
    Voir sur GitHub↗6,093
  1. Home
  2. Software Engineering & Architecture
  3. Context-Aware Tooling
  4. Visual Context Awareness Engines

Explorer les sous-tags

  • Desktop Context-Aware ShortcutsTriggers pre-configured actions or agent workflows based on the active application and visible content on the desktop. **Distinct from Visual Context Awareness Engines:** Distinct from Visual Context Awareness Engines: focuses on triggering pre-configured shortcuts and workflows rather than extracting visual metadata for LLM situational awareness.
  • URL-Based Context ResolversSystems that trigger specific AI behaviors based on the current browser URL and site configuration. **Distinct from Visual Context Awareness Engines:** Focuses on URL-based trigger mappings for browser extensions, whereas Visual Context Awareness Engines extract desktop visual metadata.
  • Web Content AwarenessSystems that analyze active web page metadata to provide context to AI models. **Distinct from Visual Context Awareness Engines:** Focuses on web page metadata rather than visual desktop coordinates.