13 مستودعات
Search systems that index technical content for machine-readable retrieval and natural language querying.
Distinct from AI-Powered Search: Distinct from general AI search: specifically optimized for technical documentation indexing and retrieval.
Explore 13 awesome GitHub repositories matching artificial intelligence & ml · Documentation Search. Refine with filters or upvote what's useful.
Context Hub is a retrieval-augmented generation framework and context management system designed to provide large language model agents with curated, versioned markdown documentation. It functions as a documentation provider that delivers precise API references and technical context to reduce hallucinations and token waste. The system incorporates an agentic memory layer that maintains persistent local annotations and user feedback to improve how agents retrieve task-specific knowledge. It uses a version-controlled repository of technical documentation designed for both machine readability an
Provides a search system optimized for indexing technical content and locating specific skills within the documentation catalog.
Fumadocs is a documentation framework designed for building content-heavy technical websites using MDX. It functions as a static site generator that transforms structured text files into optimized, interactive web pages, providing a comprehensive toolset for managing technical content, API references, and versioned guides. The platform distinguishes itself through a deep integration of interactive components and AI-ready features. It includes a library of pre-built interface elements that allow developers to embed live API playgrounds, request snippets, and schema-based documentation directly
Structures technical content for machine-readable indexing and embedding chat interfaces to help users find answers using language models.
This project is a framework for developing multimodal AI agents that function as programmable participants in real-time communication rooms. It enables the construction of agents that can see, hear, and speak by integrating speech-to-text, large language models, and text-to-speech pipelines to facilitate low-latency, natural conversations. The system is distinguished by its advanced orchestration of real-time media and conversational flow, including support for full-duplex speech, preemptive response generation, and sophisticated interruption management. It further differentiates itself throu
Enables AI agents to retrieve technical context by searching through project documentation and source code.
git-mcp is a Model Context Protocol server that transforms Git repositories and static sites into structured context providers for AI assistants. It functions as a documentation retrieval tool and repository indexer, exposing codebases and project files as standardized tools to reduce hallucinations in large language model responses. The project converts raw repository files, READMEs, and external URLs into formats optimized for token consumption. It enables AI agents to perform query-based code searches and retrieve specific sections of project documentation to maintain up-to-date technical
Implements a search system that indexes technical documentation for machine-readable retrieval and natural language querying via MCP.
Riot هو محرك بحث موزع وخادم فهرسة قائم على Go مصمم للفهرسة والاسترجاع بالنص الكامل. يعمل كنظام استرجاع يقوم بفرز المستندات حسب الصلة باستخدام خوارزميات ترتيب BM25، وتكرار المصطلح، وتكرار المستند العكسي. يوفر المحرك دعماً متخصصاً للغة الصينية، ويتميز بتجزئة النص المتزامنة وتعيين Pinyin الصوتي لمطابقة المدخلات الرومانية مع الأحرف. يستخدم بنية موزعة توظف تقسيم الفهرس القائم على التجزئة (hash-based sharding) لموازنة تحميل البيانات والإنتاجية عبر عقد خادم متعددة. يغطي النظام نطاقاً واسعاً من قدرات البحث، بما في ذلك تنفيذ استعلام المنطق البولياني، وتصفية القرب، وإدارة دورة حياة الفهرس في الوقت الفعلي. يحتفظ بفهرس سريع قابل للبحث في الذاكرة مع استخدام تخزين مدعوم بالقرص لاستمرارية البيانات ومتانتها. يتم تضمين أدوات مراقبة لتتبع استخدام الذاكرة والقرص و CPU عبر البيئة الموزعة.
Supports real-time addition and removal of documents while the engine is running to maintain continuous availability.
Enables AI assistants to search documentation and browse code examples through the Model Context Protocol.
Lingui is a JavaScript internationalization library that provides a framework-agnostic core with bindings for React, SolidJS, Svelte, Astro, and other JavaScript frameworks. It operates through a compile-time message extraction pipeline that scans source files for translatable strings, generates standard PO, JSON, or CSV catalog files, and compiles them into optimized JavaScript modules for production deployment. The library uses macro-based message definition to wrap translatable text in source code while preserving context for extraction, and includes a plural rule engine that automatically
Indexes the latest library documentation so AI agents can search it on demand with low token usage.
Potpie is an LLM codebase analysis platform and multi-agent orchestration framework designed to act as an AI software engineer. It parses repositories into a structured code knowledge graph, enabling AI agents to perform multi-hop reasoning, dependency tracing, and grounded technical analysis across large codebases. The system distinguishes itself through a spec-driven development framework where agents generate detailed technical specifications and architecture plans before implementing multi-file code changes. It utilizes a durable execution engine to coordinate specialized AI personas for
Indexes and retrieves technical content from the web to supplement codebase knowledge.
Wukong هو محرك بحث نصي كامل موزع مصمم لفهرسة واسترجاع المستندات النصية. يعمل كخلفية بحث قابلة للتخصيص تستخدم مصنف أهمية BM25 لترتيب نتائج البحث بناءً على تكرار المصطلح وتكرار المستند العكسي. يتضمن النظام مقسماً نصياً صينياً متخصصاً لتقسيم سلاسل الأحرف المستمرة إلى كلمات ذات معنى للفهرسة والاسترجاع الدقيق. للتعامل مع مجموعات البيانات الكبيرة وأحجام الطلبات العالية، يستخدم فهرس بحث موزعاً يستخدم التجزئة القائمة على الهاش لتقسيم المستندات عبر عقد متعددة. يوفر المحرك قدرات استرجاع معلومات شاملة، بما في ذلك فهرسة الرموز المدركة للموقع لتصفية القرب وترتيب النتائج متعدد المعايير. ويدعم إدارة الفهرس في الوقت الفعلي من خلال تحديثات مجمعة ويضمن توفر البيانات عبر عمليات إعادة التشغيل من خلال استمرارية بيانات الفهرس.
Updates searchable text indices in real time using batched operations to keep results current.
This project is a collection of structured study notes and conceptual breakdowns designed for the AWS Certified Cloud Practitioner exam. It serves as a technical reference and study guide, organizing cloud service details and architectural principles to assist in certification preparation. The knowledge base is built using markdown files and includes curated cheat sheets and interactive mind-map visualizations. These tools map complex certification topics into visual hierarchies to enable drill-down study paths and rapid revision. The materials cover a wide range of cloud capabilities, inclu
Outlines capabilities for indexing technical content and retrieving answers through natural language documentation search.
Geist is an open-source font family and typography collection designed for high legibility in technical interfaces. It consists of a series of web-optimized typefaces, including geometric sans-serif, monospaced, and pixel styles. The collection functions as a variable font library, utilizing coordinate interpolation to allow precise control over weight and style within a single font file. These fonts are built as OpenType typefaces, incorporating standardized layout tables to define advanced typographic behaviors such as kerning and ligatures. The project provides specific implementations fo
Indexes technical documentation to provide AI agents with necessary development context.
This project is a Model Context Protocol server and AI agent database connector. It provides a standardized communication layer that allows language models to interact with relational data stores, read database schemas, and manage PostgreSQL database resources. The implementation acts as a serverless host for the Model Context Protocol, deploying on distributed edge functions to connect AI assistants to a project. This enables AI agents to perform database administration, execute SQL queries, and handle schema migrations through an AI-compatible interface. The system covers broader capabilit
Enables AI assistants to query a technical knowledge base to retrieve context on feature usage.
مولد الخلاصات (feed generator) هو إطار عمل لبناء ونشر خلاصات محتوى خوارزمية مخصصة داخل شبكة AT Protocol. يوفر البنية التحتية لتحديد منطق تنسيق فريد، وتسجيل هذه الخوارزميات في ملفات تعريف المستخدمين، وتقديم تدفقات محتوى مخصصة للشبكة. يتميز إطار العمل بدمج فهرسة نشاط الشبكة في الوقت الفعلي مع نظام توجيه قائم على المعالجات (handler-based). ومن خلال استهلاك تدفقات الأحداث المباشرة، فإنه يحتفظ بمجموعات بيانات محلية تسمح للمطورين بتطبيق قواعد فرز وتصفية مخصصة على المحتوى العام. كما يدير دورة حياة هذه الخلاصات من خلال التسجيل التعريفي، مما يمكن ربط الخدمات بمالكي الحسابات واكتشافها عبر النظام البيئي اللامركزي. يتضمن النظام أدوات أمنية وتشغيلية مدمجة للتعامل مع الطلبات الموثقة واسترجاع البيانات على نطاق واسع. ويستخدم التحقق من الهوية القائم على الرموز (token-based) لضمان الوصول الآمن إلى بيانات المستخدم المحددة، ويستخدم الترقيم القائم على المؤشر (cursor-based) لإدارة استخدام الذاكرة عند تقديم مجموعات نتائج كبيرة. تتم إدارة التهيئة من خلال متغيرات البيئة الخارجية، مما يسمح بإجراء تعديلات على سلوك الخدمة دون الحاجة إلى إعادة نشر الكود.
Consumes live event streams to build local databases for content filtering.