awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

5 مستودعات

Awesome GitHub RepositoriesVisual Context Awareness Engines

Systems that extract visual metadata from the desktop to provide situational awareness to language models.

Distinct from Context-Aware Tooling: Distinct from general context-aware tooling: focuses on visual desktop metadata extraction for LLM situational awareness.

Explore 5 awesome GitHub repositories matching software engineering & architecture · Visual Context Awareness Engines. Refine with filters or upvote what's useful.

Awesome Visual Context Awareness Engines GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • moeru-ai/airiالصورة الرمزية لـ moeru-ai

    moeru-ai/airi

    41,040عرض على GitHub↗

    Airi is an interactive digital companion engine designed to bridge large language models with local animation rendering. It functions as a middleware platform that synchronizes conversational text streams with skeletal and facial movements to drive virtual avatars in real time. The framework distinguishes itself by integrating desktop context awareness, allowing characters to maintain situational awareness of a user's screen activity across both desktop and web environments. It utilizes a hybrid execution model that splits computational workloads between cloud-based language processing and lo

    Extracts visual metadata from the user environment to provide language models with situational awareness of the active desktop.

    TypeScriptai-companionai-vtuberclawdbot
    عرض على GitHub↗41,040
  • zuodaotech/everyone-can-use-englishالصورة الرمزية لـ ZuodaoTech

    ZuodaoTech/everyone-can-use-english

    34,948عرض على GitHub↗

    This project is an AI-powered English education tool and browser extension designed for immersive language learning. It functions as a proficiency training suite that integrates AI tools and linguistic analysis directly into external websites and streaming video platforms. The system employs a DOM injection model to add interactive overlays and toolbars to web pages. It uses large language model APIs to provide real-time translations and maps curated pronunciation and grammar exercises to specific timestamps in external media. The tool covers a broad range of English proficiency training, in

    Uses page metadata and active media content to provide real-time translations and linguistic analysis via LLMs.

    TypeScript
    عرض على GitHub↗34,948
  • mediar-ai/screenpipeالصورة الرمزية لـ mediar-ai

    mediar-ai/screenpipe

    19,337عرض على GitHub↗

    Screenpipe is a local screen and audio recorder that captures and indexes digital activity to create a searchable archive of computer usage. It functions as an AI context engine, providing a local database of visual and auditory history to ground large language models. The system serves as a Model Context Protocol server, delivering screen history and meeting transcriptions to external AI assistants. It utilizes an OCR screen search tool to extract text from visual data and a speech-to-text transcription tool for identifying speakers in system and microphone audio. The software includes capa

    Functions as a visual context awareness engine that grounds LLMs with a local database of visual and auditory history.

    Rust
    عرض على GitHub↗19,337
  • josstorer/chatgptboxالصورة الرمزية لـ josStorer

    josStorer/chatGPTBox

    10,738عرض على GitHub↗

    chatGPTBox is a browser extension that integrates large language model chat interfaces and AI tools directly into the web browsing experience. It functions as an AI productivity toolkit and API client, allowing users to access AI assistants via a floating chat interface without leaving their active webpage. The project distinguishes itself by offering context-aware assistance and website-specific adaptations based on the current URL. It further enhances the browsing experience by displaying AI-generated responses alongside standard search engine results and providing a system to route chat re

    Uses site-specific configuration mappings to trigger tailored AI behaviors based on the current webpage URL.

    JavaScript
    عرض على GitHub↗10,738
  • sylinko/everywhereالصورة الرمزية لـ Sylinko

    Sylinko/Everywhere

    6,093عرض على GitHub↗

    Everywhere is a desktop AI assistant that understands whatever is on your screen and can act across applications without requiring screenshots or manual context switching. It reads structured UI data through accessibility and automation APIs to perceive the active application and visible content, then provides context-aware help, summaries, translations, and answers to natural language questions about what you are viewing. The tool distinguishes itself by combining on-screen content analysis with a multi-LLM agent platform that routes requests to providers like OpenAI, Anthropic, and local mo

    Triggers pre-configured actions or agent workflows based on the active application and visible content.

    C#aiai-agentsai-assistant
    عرض على GitHub↗6,093
  1. Home
  2. Software Engineering & Architecture
  3. Context-Aware Tooling
  4. Visual Context Awareness Engines

استكشف الوسوم الفرعية

  • Desktop Context-Aware ShortcutsTriggers pre-configured actions or agent workflows based on the active application and visible content on the desktop. **Distinct from Visual Context Awareness Engines:** Distinct from Visual Context Awareness Engines: focuses on triggering pre-configured shortcuts and workflows rather than extracting visual metadata for LLM situational awareness.
  • URL-Based Context ResolversSystems that trigger specific AI behaviors based on the current browser URL and site configuration. **Distinct from Visual Context Awareness Engines:** Focuses on URL-based trigger mappings for browser extensions, whereas Visual Context Awareness Engines extract desktop visual metadata.
  • Web Content AwarenessSystems that analyze active web page metadata to provide context to AI models. **Distinct from Visual Context Awareness Engines:** Focuses on web page metadata rather than visual desktop coordinates.