awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 Repos

Awesome GitHub RepositoriesVisual Context Awareness Engines

Systems that extract visual metadata from the desktop to provide situational awareness to language models.

Distinct from Context-Aware Tooling: Distinct from general context-aware tooling: focuses on visual desktop metadata extraction for LLM situational awareness.

Explore 5 awesome GitHub repositories matching software engineering & architecture · Visual Context Awareness Engines. Refine with filters or upvote what's useful.

Awesome Visual Context Awareness Engines GitHub Repositories

Finde die besten Repos mit KI.Wir suchen mit KI nach den am besten passenden Repositories.
  • moeru-ai/airiAvatar von moeru-ai

    moeru-ai/airi

    41,040Auf GitHub ansehen↗

    Airi is an interactive digital companion engine designed to bridge large language models with local animation rendering. It functions as a middleware platform that synchronizes conversational text streams with skeletal and facial movements to drive virtual avatars in real time. The framework distinguishes itself by integrating desktop context awareness, allowing characters to maintain situational awareness of a user's screen activity across both desktop and web environments. It utilizes a hybrid execution model that splits computational workloads between cloud-based language processing and lo

    Extracts visual metadata from the user environment to provide language models with situational awareness of the active desktop.

    TypeScriptai-companionai-vtuberclawdbot
    Auf GitHub ansehen↗41,040
  • zuodaotech/everyone-can-use-englishAvatar von ZuodaoTech

    ZuodaoTech/everyone-can-use-english

    34,948Auf GitHub ansehen↗

    This project is an AI-powered English education tool and browser extension designed for immersive language learning. It functions as a proficiency training suite that integrates AI tools and linguistic analysis directly into external websites and streaming video platforms. The system employs a DOM injection model to add interactive overlays and toolbars to web pages. It uses large language model APIs to provide real-time translations and maps curated pronunciation and grammar exercises to specific timestamps in external media. The tool covers a broad range of English proficiency training, in

    Uses page metadata and active media content to provide real-time translations and linguistic analysis via LLMs.

    TypeScript
    Auf GitHub ansehen↗34,948
  • mediar-ai/screenpipeAvatar von mediar-ai

    mediar-ai/screenpipe

    19,337Auf GitHub ansehen↗

    Screenpipe is a local screen and audio recorder that captures and indexes digital activity to create a searchable archive of computer usage. It functions as an AI context engine, providing a local database of visual and auditory history to ground large language models. The system serves as a Model Context Protocol server, delivering screen history and meeting transcriptions to external AI assistants. It utilizes an OCR screen search tool to extract text from visual data and a speech-to-text transcription tool for identifying speakers in system and microphone audio. The software includes capa

    Functions as a visual context awareness engine that grounds LLMs with a local database of visual and auditory history.

    Rust
    Auf GitHub ansehen↗19,337
  • josstorer/chatgptboxAvatar von josStorer

    josStorer/chatGPTBox

    10,738Auf GitHub ansehen↗

    chatGPTBox is a browser extension that integrates large language model chat interfaces and AI tools directly into the web browsing experience. It functions as an AI productivity toolkit and API client, allowing users to access AI assistants via a floating chat interface without leaving their active webpage. The project distinguishes itself by offering context-aware assistance and website-specific adaptations based on the current URL. It further enhances the browsing experience by displaying AI-generated responses alongside standard search engine results and providing a system to route chat re

    Uses site-specific configuration mappings to trigger tailored AI behaviors based on the current webpage URL.

    JavaScript
    Auf GitHub ansehen↗10,738
  • sylinko/everywhereAvatar von Sylinko

    Sylinko/Everywhere

    6,093Auf GitHub ansehen↗

    Everywhere is a desktop AI assistant that understands whatever is on your screen and can act across applications without requiring screenshots or manual context switching. It reads structured UI data through accessibility and automation APIs to perceive the active application and visible content, then provides context-aware help, summaries, translations, and answers to natural language questions about what you are viewing. The tool distinguishes itself by combining on-screen content analysis with a multi-LLM agent platform that routes requests to providers like OpenAI, Anthropic, and local mo

    Triggers pre-configured actions or agent workflows based on the active application and visible content.

    C#aiai-agentsai-assistant
    Auf GitHub ansehen↗6,093
  1. Home
  2. Software Engineering & Architecture
  3. Context-Aware Tooling
  4. Visual Context Awareness Engines

Unter-Tags erkunden

  • Desktop Context-Aware ShortcutsTriggers pre-configured actions or agent workflows based on the active application and visible content on the desktop. **Distinct from Visual Context Awareness Engines:** Distinct from Visual Context Awareness Engines: focuses on triggering pre-configured shortcuts and workflows rather than extracting visual metadata for LLM situational awareness.
  • URL-Based Context ResolversSystems that trigger specific AI behaviors based on the current browser URL and site configuration. **Distinct from Visual Context Awareness Engines:** Focuses on URL-based trigger mappings for browser extensions, whereas Visual Context Awareness Engines extract desktop visual metadata.
  • Web Content AwarenessSystems that analyze active web page metadata to provide context to AI models. **Distinct from Visual Context Awareness Engines:** Focuses on web page metadata rather than visual desktop coordinates.