21 Repos
Mechanisms for fetching detailed operational instructions only when a specific capability is triggered.
Distinct from Project Context Managers: Focuses on token-reduction via lazy-loading of instructions rather than static project constraints.
Explore 21 awesome GitHub repositories matching software engineering & architecture · On-Demand Context Loading. Refine with filters or upvote what's useful.
This project is a structured educational resource and technical guide for designing and implementing autonomous systems using large language models. It provides a comprehensive curriculum and code samples focused on agentic design patterns, autonomous development, and the creation of systems capable of planning and executing multi-step tasks. The resource details the implementation of agentic retrieval-augmented generation, where models autonomously plan and refine data searches. It covers a wide array of orchestrators and design patterns, including metacognitive reflection for self-correctin
Injects relevant background information into the model before processing a query to ensure informed responses.
Agent-skills is a collection of structured instructions and behavioral personas designed to standardize how AI coding agents perform engineering tasks. It functions as a workflow orchestrator that maps natural language intent to repeatable technical sequences and verification checklists. The project distinguishes itself through the use of specialized markdown-defined roles, such as security auditors or test engineers, to apply targeted domain expertise. It employs an evidence-based verification model that requires runtime data or passing tests as mandatory exit criteria to ensure AI-generated
Allows for the explicit referencing of skill files within prompts to force specific workflows on-demand.
vim-plug is a plugin manager for the Vim text editor that functions as a git-based dependency manager and configuration bootstrapper. It downloads and organizes external plugin packages from remote repositories to extend the editor's functionality. The project acts as a lazy-loading orchestrator to reduce editor startup time by deferring the loading of plugins until specific commands or file types are encountered. It enables version pinning by allowing plugins to be fetched using specific git branches, tags, or commit hashes. The manager provides a framework for plugin installation, lifecycl
Reduces editor startup time by loading plugins only when specific commands or file types are encountered.
SpaceVim is a modular configuration framework for Vim and Neovim designed to manage settings, plugins, and keybindings across different editing environments. It functions as a plugin manager that uses a layered organization system to group related functions and plugins, reducing the manual effort required for configuration. The system is centered around a mnemonic keybinding strategy that utilizes the space bar as the primary modifier for editor commands. To maintain performance, the framework implements a startup optimizer that delays the loading of non-essential plugins until they are requi
Implements a deferred plugin loading mechanism to minimize initial editor startup time.
SpaceVim is a modular configuration framework designed for managing plugins and settings across Vim and Neovim. It utilizes a layer-based architecture to organize related editor extensions into functional groups, reducing the overhead associated with manual setup and maintenance. The framework features a mnemonic keybinding system that maps editor commands to intuitive key sequences and provides integrated discovery guides. To improve performance, it employs a deferred plugin loader that reduces startup time by loading extensions only when they are specifically required. The system provides
Implements a deferred plugin loader that increases startup speed by delaying initialization until modules are requested.
Hermes-webui is a self-hosted AI orchestrator and web interface for managing autonomous agents. It serves as a multi-provider gateway that connects cloud and local large language models, providing a central hub to execute scheduled background jobs, run shell commands, and manage agent memory on private hardware. The system distinguishes itself through a persistent memory manager that utilizes knowledge graphs and markdown files for long-term context across sessions. It features a model context protocol host for extending agent capabilities with standardized tools and supports the orchestratio
Writes new skill modules to extend functional abilities based on learned context and operational experience.
OpenHarness is a framework for building and orchestrating AI agents that utilize tools and plugins to execute complex tasks. It provides an orchestration system for managing language model lifecycles and a multi-agent coordination system for delegating workloads across teams of specialized subagents. The project features an agent gateway that bridges language model agents to external chat platforms and communication channels. It includes a tool integration engine for executing shell, file, and web operations, supported by a memory and skill manager that handles persistent user preferences and
Implements modular, domain-specific instruction files loaded into the agent context to guide task execution.
Context Hub is a retrieval-augmented generation framework and context management system designed to provide large language model agents with curated, versioned markdown documentation. It functions as a documentation provider that delivers precise API references and technical context to reduce hallucinations and token waste. The system incorporates an agentic memory layer that maintains persistent local annotations and user feedback to improve how agents retrieve task-specific knowledge. It uses a version-controlled repository of technical documentation designed for both machine readability an
Retrieves only the minimal subset of required reference files based on current task requirements to optimize tokens.
Agent Skills is a framework for bundling executable scripts and metadata to extend the capabilities and tool-use of language model agents. It provides a standardized directory structure for packaging specialized workflows, technical instructions, and portable agent capabilities for distribution across different AI platforms. The project features a tool optimization suite used to refine skill triggers and evaluate the reliability of agent-activated capabilities. It includes a context-aware knowledge manager that organizes technical references into a hierarchy, loading them on demand to reduce
Reduces token consumption by loading full technical instructions on demand rather than during initial discovery.
This project is an AI agent workflow orchestrator and automated software lifecycle manager designed to sequence specialized AI personas for end-to-end software development. It serves as a prompt engineering library and a full-stack development toolkit that guides the process from initial discovery and specification through to deployment and code review. The system features a context management framework that utilizes progressive loading and routing tables to fetch reference files on-demand, reducing token consumption within the model context window. It employs a definition-based routing syste
Reduces token consumption by lazily fetching detailed reference data and documentation only when specific triggers occur.
mini.nvim is a comprehensive library of independent modules designed to extend Neovim with a wide array of navigation, user interface, and text manipulation tools. It serves as a modular plugin collection, a UI toolkit for creating custom statuslines and notifications, and a package manager for installing and pinning external plugins from Git. The project provides a specialized fuzzy picker framework for filtering files and symbols, an LSP completion engine with interactive snippet expansion, and a dedicated plugin test framework that uses headless editor instances and remote procedure calls
Optimizes startup time by scheduling the execution of plugin setup code to occur after the initial load.
Evolver is a self-evolving AI agent framework that uses gene expression programming to autonomously improve agent behaviors through a continuous five-step loop of scanning, selecting, mutating, validating, and solidifying. It functions as an auditable evolution system that records every mutation and selection step, and can translate natural-language problems into executable Python code for automated grading and evaluation. The framework distinguishes itself through a distributed architecture that enables multiple agents to collaborate and share learned experiences across a network. It operate
Selects relevant Genes or Capsules at runtime by computing signal overlap and loads their full content only when needed.
Feast is an open-source feature store for machine learning that provides a central platform for defining, storing, and serving features across both training and inference workflows. It operates as a declarative system where feature definitions are written as code in Python files, synchronized to a central registry, and made available for low-latency online retrieval or point-in-time correct historical joins for training datasets. The project abstracts storage behind a pluggable architecture, allowing offline and online backends to be swapped without changing retrieval logic, and coordinates ma
Pre-loads heavy resources like models and lookup tables at server startup to reduce per-request overhead.
zplug ist ein Plugin-Manager für die Zsh-Shell, der dazu entwickelt wurde, Community-Plugins, Themes und Befehle aus entfernten Repositories und lokalen Pfaden zu installieren und zu organisieren. Er konzentriert sich auf die Reproduzierbarkeit der Umgebung durch versionierte Konfiguration, was es ermöglicht, Plugins an spezifische Git-Branches, Tags oder Commit-Hashes zu pinnen. Das Projekt optimiert die Shell-Startgeschwindigkeit durch ein Lazy-Loading-System, das die Ausführung von Plugins verzögert, bis sie explizit benötigt werden. Um die anfängliche Umgebungseinrichtung zu beschleunigen, setzt es einen parallelen Installer ein, der mehrere Erweiterungen gleichzeitig herunterlädt. Der Manager unterstützt Multi-Source-Installation und handhabt die Abhängigkeitsauflösung von Erweiterungen, um sicherzustellen, dass erforderliche Pakete in der richtigen Reihenfolge geladen werden. Er enthält zudem Funktionen zum Herunterladen externer binärer Artefakte und deren automatische Zuordnung zum Systempfad sowie zur Ausführung benutzerdefinierter Lebenszyklus-Hooks während der Installation und Updates.
Implements a system to delay the activation of plugins until a specific command is first invoked, optimizing startup speed.
Perfetto is a platform for system-level performance tracing and analysis on Linux and Android. It combines a high-throughput trace recorder, a SQL-based query engine, and a browser-based visualizer into a single toolchain. The platform covers CPU scheduling and call-stack profiling, native and Java heap memory allocation tracking, GPU and graphics events, and system-wide counters such as CPU frequency and power consumption. The architecture decouples trace recording from offline analysis, using a compact protobuf format for event encoding and columnar storage for efficient SQL queries. The we
Provides a library of curated query examples that users can load for trace analysis.
Copilot ist eine selbst gehostete Kundensupport-Plattform, die Large Language Models und vektorisierte Wissensdatenbanken verwendet, um die Benutzerunterstützung zu automatisieren. Das System besteht aus einem einbettbaren KI-Chat-Widget für die Website-Integration, einem containerisierten Management-Dashboard und einer Vektor-Datenbank als Wissensbasis. Die Plattform enthält einen API-Orchestrator, der strukturierte API-Spezifikationen verarbeitet und es Sprachmodellen ermöglicht, mit externen Endpunkten und Operationen zu interagieren. Sie verwaltet spezialisierten Kontext für Antworten durch die Indexierung von API-Definitionen und Dokumentationen innerhalb einer Vektor-Datenbank. Die Infrastruktur deckt die containerisierte Service-Orchestrierung für Server und Hintergrund-Worker, asynchrone Queue-basierte Verarbeitung sowie ein administratives Interface zur Überwachung von Echtzeit-Benutzerinteraktionen ab. Das System handhabt zudem Datenbank-Schema-Migrationen und Wissensindexierung.
Injects structured API definitions as context into the language model to enable automated interaction with external operations.
Antigravity-kit is a multi-agent orchestrator and routing engine designed to coordinate specialized large language model agents. It functions as a conversational workflow automation tool and a context management system that executes complex tasks through a chat interface. The system utilizes a routing engine to classify user requests and dispatch them to domain-expert agents. It employs a multi-agent orchestration model that allows specialist workers to operate in parallel and combine their outputs. To manage operational efficiency, the kit includes a memory layer for storing project convent
Fetches detailed operational instructions only when specific capabilities are triggered to reduce token overhead.
Kiln ist eine LLM-Entwicklungs-Workbench und ein Evaluierungs-Framework, das für das Design, das Testen und die Optimierung von Prompts und KI-Agenten entwickelt wurde. Es fungiert als Multi-Agenten-Orchestrator und RAG-Optimierungstool und bietet eine visuelle Schnittstelle für die iterative Entwicklung von KI-Systemen. Das Projekt zeichnet sich durch eine umfassende Fine-Tuning-Pipeline aus, die Zero-Code-Modelltraining und Reasoning-Destillation unterstützt. Es ermöglicht die Erstellung hierarchischer Multi-Agenten-Systeme, in denen spezialisierte Akteure über Tool-Calling koordinieren, und implementiert einen Model Context Protocol-Server, um diese Agenten und Suchfunktionen als standardisierte Tools für externe Clients bereitzustellen. Die Plattform deckt ein breites Spektrum an Fähigkeiten ab, einschließlich automatisierter KI-Jury-Bewertung zur Qualitätssicherung, synthetischer Datengenerierung für Training und Evaluierung sowie hybrider Vektor-Keyword-Suche zur Erdung von Modellantworten. Sie bietet zudem Tools für Prompt-Evolution, Trace-Auditing und die Verwaltung kollaborativer Datensätze via Git-Integration. Die Workbench ist über eine selbst-hostbare REST-API und eine dedizierte Python-Bibliothek für die programmatische Workflow-Ausführung zugänglich.
Defines reusable instructions and reference documents that are loaded into agent context only when needed.
Youtu Agent is an open-source framework for building, running, and evaluating autonomous agents powered by large language models. It provides the core infrastructure for creating agents that follow reasoning loops, use toolkits, and coordinate with other agents to solve complex tasks, all managed through YAML-driven configuration files. The framework distinguishes itself through its support for multi-agent orchestration, where a planner agent decomposes tasks and coordinates specialized worker agents, and through its integration with the Model Context Protocol for connecting to external toolk
Provides modular skill files that extend agent capabilities with structured instructions for complex tasks.
Dieses Projekt ist ein deklaratives Framework und eine domänenspezifische Sprache zur Verwaltung von Emacs-Lisp-Paketen. Es fungiert als Optimierer für die Start-Performance, indem es Paketinstallation, Variableneinstellungen und Tastenkombinationen in einzelnen Blöcken gruppiert, um die anfängliche Boot-Zeit zu reduzieren. Das System zeichnet sich durch ein Framework für verzögertes Laden aus, das die Paketausführung verzögert, bis spezifische Tasten, Hooks oder Modi ausgelöst werden. Es verwendet eine makrobasierte Deklarationssyntax, um die Konfiguration zu organisieren und die Generierung von Autoloads zu automatisieren, wodurch sichergestellt wird, dass Pakete nur geladen werden, wenn sie tatsächlich benötigt werden. Das Projekt deckt breite Funktionsbereiche ab, einschließlich automatisierter Paketinstallation aus Remote-Archiven, Abhängigkeitsmanagement und bedingtem Laden basierend auf dem Betriebssystem oder der Umgebung. Es bietet zudem umfassende Tastaturbelegung für Shortcuts und Chords sowie Tools zur Analyse von Ladestatistiken und zur Modifikation von Mode-Line-Anzeigen.
Provides a system that delays the activation of packages until a specific command, hook, or mode is triggered.