awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 repositorios

Awesome GitHub RepositoriesVisual Context Awareness Engines

Systems that extract visual metadata from the desktop to provide situational awareness to language models.

Distinct from Context-Aware Tooling: Distinct from general context-aware tooling: focuses on visual desktop metadata extraction for LLM situational awareness.

Explore 5 awesome GitHub repositories matching software engineering & architecture · Visual Context Awareness Engines. Refine with filters or upvote what's useful.

Awesome Visual Context Awareness Engines GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • moeru-ai/airiAvatar de moeru-ai

    moeru-ai/airi

    41,040Ver en GitHub↗

    Airi is an interactive digital companion engine designed to bridge large language models with local animation rendering. It functions as a middleware platform that synchronizes conversational text streams with skeletal and facial movements to drive virtual avatars in real time. The framework distinguishes itself by integrating desktop context awareness, allowing characters to maintain situational awareness of a user's screen activity across both desktop and web environments. It utilizes a hybrid execution model that splits computational workloads between cloud-based language processing and lo

    Extracts visual metadata from the user environment to provide language models with situational awareness of the active desktop.

    TypeScriptai-companionai-vtuberclawdbot
    Ver en GitHub↗41,040
  • zuodaotech/everyone-can-use-englishAvatar de ZuodaoTech

    ZuodaoTech/everyone-can-use-english

    34,948Ver en GitHub↗

    This project is an AI-powered English education tool and browser extension designed for immersive language learning. It functions as a proficiency training suite that integrates AI tools and linguistic analysis directly into external websites and streaming video platforms. The system employs a DOM injection model to add interactive overlays and toolbars to web pages. It uses large language model APIs to provide real-time translations and maps curated pronunciation and grammar exercises to specific timestamps in external media. The tool covers a broad range of English proficiency training, in

    Uses page metadata and active media content to provide real-time translations and linguistic analysis via LLMs.

    TypeScript
    Ver en GitHub↗34,948
  • mediar-ai/screenpipeAvatar de mediar-ai

    mediar-ai/screenpipe

    19,337Ver en GitHub↗

    Screenpipe is a local screen and audio recorder that captures and indexes digital activity to create a searchable archive of computer usage. It functions as an AI context engine, providing a local database of visual and auditory history to ground large language models. The system serves as a Model Context Protocol server, delivering screen history and meeting transcriptions to external AI assistants. It utilizes an OCR screen search tool to extract text from visual data and a speech-to-text transcription tool for identifying speakers in system and microphone audio. The software includes capa

    Functions as a visual context awareness engine that grounds LLMs with a local database of visual and auditory history.

    Rust
    Ver en GitHub↗19,337
  • josstorer/chatgptboxAvatar de josStorer

    josStorer/chatGPTBox

    10,738Ver en GitHub↗

    chatGPTBox is a browser extension that integrates large language model chat interfaces and AI tools directly into the web browsing experience. It functions as an AI productivity toolkit and API client, allowing users to access AI assistants via a floating chat interface without leaving their active webpage. The project distinguishes itself by offering context-aware assistance and website-specific adaptations based on the current URL. It further enhances the browsing experience by displaying AI-generated responses alongside standard search engine results and providing a system to route chat re

    Uses site-specific configuration mappings to trigger tailored AI behaviors based on the current webpage URL.

    JavaScript
    Ver en GitHub↗10,738
  • sylinko/everywhereAvatar de Sylinko

    Sylinko/Everywhere

    6,093Ver en GitHub↗

    Everywhere is a desktop AI assistant that understands whatever is on your screen and can act across applications without requiring screenshots or manual context switching. It reads structured UI data through accessibility and automation APIs to perceive the active application and visible content, then provides context-aware help, summaries, translations, and answers to natural language questions about what you are viewing. The tool distinguishes itself by combining on-screen content analysis with a multi-LLM agent platform that routes requests to providers like OpenAI, Anthropic, and local mo

    Triggers pre-configured actions or agent workflows based on the active application and visible content.

    C#aiai-agentsai-assistant
    Ver en GitHub↗6,093
  1. Home
  2. Software Engineering & Architecture
  3. Context-Aware Tooling
  4. Visual Context Awareness Engines

Explorar subetiquetas

  • Desktop Context-Aware ShortcutsTriggers pre-configured actions or agent workflows based on the active application and visible content on the desktop. **Distinct from Visual Context Awareness Engines:** Distinct from Visual Context Awareness Engines: focuses on triggering pre-configured shortcuts and workflows rather than extracting visual metadata for LLM situational awareness.
  • URL-Based Context ResolversSystems that trigger specific AI behaviors based on the current browser URL and site configuration. **Distinct from Visual Context Awareness Engines:** Focuses on URL-based trigger mappings for browser extensions, whereas Visual Context Awareness Engines extract desktop visual metadata.
  • Web Content AwarenessSystems that analyze active web page metadata to provide context to AI models. **Distinct from Visual Context Awareness Engines:** Focuses on web page metadata rather than visual desktop coordinates.