25 مستودعات
Standardized APIs that provide a consistent execution interface across different language model providers.
Distinct from Unified Model Wrappers: Distinct from wrappers as it focuses on the standardized execution interface for processing and streaming across providers.
Explore 25 awesome GitHub repositories matching software engineering & architecture · Unified Model Interfaces. Refine with filters or upvote what's useful.
Pi is an autonomous coding agent and framework for building AI agents capable of executing independent loops. It functions as an agent state management system that tracks and persists tool calls throughout complex workflows, utilizing a command-line interface for interaction and control. The system features a self-extensible design, allowing agents to write and implement new capabilities and tools into their own runtime environment. It also includes a provider-agnostic abstraction layer that standardizes interactions across different large language model providers through a unified API. The
Provides a standardized API layer that ensures a consistent execution interface across different language model providers.
Hutool is a comprehensive suite of Java extensions designed to serve as a standard library extension. Its primary purpose is to reduce development boilerplate for common programming tasks and data manipulation through a collection of utility classes. The project provides specialized toolkits for database management using active record patterns and connection pooling, as well as network communication via a simplified HTTP client and asynchronous socket management. It includes security and identity capabilities such as symmetric and asymmetric encryption, image captcha generation, and JWT token
Standardizes communication with different large language model providers through a common execution interface.
Llama-stack هو مكدس تنظيم موحد وبوابة API للذكاء الاصطناعي التوليدي. يوفر طبقة اتصال موحدة وواجهة متسقة لنشر وإدارة والتفاعل مع مختلف مزودي ونماذج اللغات الكبيرة. يعمل النظام كإطار عمل للوكلاء (agent framework) يدير تنفيذ الأدوات وحزم المهارات ذات الإصدارات لأتمتة المهام المعقدة. يتضمن نظام معالجة دفعات للتعامل مع كميات كبيرة من الطلبات غير المتزامنة من خلال المعالجة دون اتصال، وواجهة قاعدة بيانات متجهة لتخزين والبحث في المستندات لتمكين التوليد المعزز بالاسترجاع (RAG). يغطي المكدس قدرات عالية المستوى بما في ذلك تنظيم وكلاء الذكاء الاصطناعي، ونشر النماذج، وتوحيد واجهات برمجة تطبيقات النماذج للسماح بالتبديل بين المزودين دون إعادة كتابة تعليمات برمجية للتطبيق.
Implements a standardized execution interface for processing and streaming across different language model providers.
Agent Squad is a multi-agent system orchestrator and language model agent orchestration framework. It serves as an AI workflow automation engine and tool integration layer designed to coordinate teams of specialized agents to solve complex tasks through routing, parallel execution, and state management. The project is distinguished by its ability to dynamically compose purpose-specific agents on-demand and route requests based on intent, language, or domain expertise. It supports advanced coordination patterns, including parallel subtask distribution, sequential task pipelines, and the abilit
Standardizes execution across different language model providers through a common API for processing and streaming.
Manifest is a language model provider unification system that standardizes access to multiple AI backends through a single interface. It functions as a centralized management layer for integrating various cloud-based and local model providers to simplify how applications request completions. The system provides intelligent model routing and high availability infrastructure by directing queries based on complexity and automatically triggering model fallbacks when a primary provider fails. It distinguishes itself through multi-tenant AI management, organizing agents into isolated groups with de
Provides a standardized API interface that abstracts diverse AI model providers into a single request format.
Guardrails is a Python SDK that wraps calls to large language models with configurable validation pipelines, corrective actions, and structured output generation. It provides a unified API layer that connects to over 100 language models, applying consistent validation, streaming, and error-handling across providers. The framework validates and corrects model responses against safety and quality rules, detecting and mitigating risks in both inputs and outputs using pre-built and custom validators. The project distinguishes itself through a validator-pipeline architecture that sequentially appl
Provides a single API pattern to call any of 100+ language models with consistent validation and error handling.
SpringBlade is a development framework and platform designed for building multi-tenant SaaS applications. It provides a comprehensive scaffold for both Spring Cloud microservices and monolithic Spring Boot architectures, enabling the rapid construction of enterprise-grade software. The platform distinguishes itself through integrated LLM orchestration and industrial IoT management. It features an LLM orchestration platform that combines large language models with knowledge bases and visual AI agent workflows, alongside an IoT hub for device connectivity, state synchronization, and edge flow o
Provides a standardized API interface to connect various AI models with smart routing and real-time streaming.
Swarms is a multi-agent orchestration framework and autonomous agent toolkit designed to coordinate large language model agents. It serves as a workflow engine for managing agent relationships, providing the infrastructure to build autonomous agents with integrated memory, tool-calling capabilities, and reasoning loops. The framework is distinguished by its multi-agent consensus systems, which utilize voting, adversarial debates, and judge agents to synthesize high-quality responses. It supports a variety of collaboration patterns, including director-worker hierarchies, expert synthesis, and
Provides a standardized API that allows swapping diverse LLM providers without changing implementation code.
This project is a multimodal AI proxy and content generation hub that provides a unified web interface for interacting with multiple large language models and generative AI services. It functions as a secure API access gateway, routing requests from a single dashboard to various external AI backends using configurable base URLs and API keys. The platform is delivered as a cross-platform progressive web application, allowing for installation on Linux, Windows, and MacOS. It distinguishes itself by consolidating text, image, audio, and video generative controls into a standardized interface, su
Provides a standardized interface for interacting with multiple large language model providers.
Wenda هي منصة لتنسيق النماذج اللغوية الكبيرة (LLM) ومحرك سير عمل مخصص مصمم لإدارة خلفيات نماذج لغوية متعددة من خلال واجهة موحدة. تعمل كبوابة ذكاء اصطناعي مستضافة ذاتياً تتيح تنفيذ تسلسلات مهام معقدة وتدفقات محادثة مؤتمتة. يستخدم النظام إضافات JavaScript لتنسيق سير العمل وتشغيل استدعاءات API الخارجية. يدعم التوليد المعزز بالاسترجاع (RAG) عن طريق حقن البيانات ذات الصلة من مخازن المتجهات والملفات غير المتصلة بالإنترنت في المطالبات لزيادة دقة الاستجابة. تم بناء المنصة لنشر الشبكات الخاصة، وتتميز بإدارة الوصول متعدد المستخدمين والقدرة على تشغيل نماذج مفتوحة المصدر مكممة لتناسب قيود أجهزة معينة. كما تتضمن تتبع التاريخ القائم على الجلسة للحفاظ على سياق المحادثة.
Provides a standardized execution interface across different language model providers for seamless switching of weights and APIs.
Genkit هو إطار عمل لتطبيقات النماذج اللغوية الكبيرة (LLM) ومجموعة أدوات مطوري الذكاء الاصطناعي التوليدي المصممة لبناء تطبيقات ذكاء اصطناعي جاهزة للإنتاج. يعمل كمنسق لسير عمل الذكاء الاصطناعي الذي ينسق استدعاءات النماذج واستخدام الأدوات الوكيلة من خلال تدفقات تنفيذ آمنة من حيث النوع (type-safe). يوفر المشروع واجهة نموذج موحدة وبنية إضافات (plugin) لتوحيد الوصول إلى نماذج لغوية كبيرة متنوعة، ومخازن المتجهات، وخلفيات القياس عن بُعد. يتميز بمجموعة مراقبة مخصصة لتتبع خطوات التنفيذ ومجموعة أدوات للمطورين للتوجيه (prompting)، وتصحيح الأخطاء، وتقييم منطق الذكاء الاصطناعي عبر واجهة محلية. يغطي إطار العمل مساحة قدرات واسعة بما في ذلك تنسيق الوكلاء مع استدعاء الأدوات وتفويض الوكلاء الفرعيين، والتوليد المعزز بالاسترجاع (RAG) عبر تكامل قاعدة بيانات المتجهات، وتوليد المخرجات المهيكلة باستخدام التحقق القائم على المخطط (schema). كما يتضمن أنظمة لإدارة الجلسات ذات الحالة، وبث الاستجابات القائم على الأحداث، والقدرة على عرض تدفقات الذكاء الاصطناعي كنقاط نهاية HTTP قابلة للتوسع. يتم دعم التطوير بواسطة واجهة سطر أوامر لتشغيل الدوال وإدارة السجلات.
Provides a standardized API that maintains a consistent execution interface across diverse model providers.
Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with large language models. By acting as a reverse-proxy, it provides a centralized layer for routing requests across multiple AI providers, allowing developers to maintain consistent application logic while gaining deep visibility into model performance, usage, and costs. The platform distinguishes itself through a robust suite of traffic management and prompt engineering tools. It enables policy-driven control, including automatic failover between providers, rate limiting, and edge-b
Standardizes reasoning parameters across different AI providers to maintain a consistent interface for developers.
TaskingAI هو منسق وكلاء ذكاء اصطناعي ومنصة تطبيقات تستخدم لبناء ونشر وتوسيع نطاق التطبيقات الأصلية للذكاء الاصطناعي. يعمل كخلفية متعددة المستأجرين كخدمة، مما يوفر البنية التحتية لاستضافة وإدارة حالات وكلاء الذكاء الاصطناعي المستقلة عبر مستخدمين أو مؤسسات متعددة على بنية مشتركة. تتميز المنصة بمنشئ سير عمل مرئي ووحدة تحكم لإدارة المشاريع، مما يسمح للمستخدمين بتهيئة منطق الوكيل واختبار سير عمل المحادثة من خلال واجهة رسومية قبل نقلها إلى بيئة الإنتاج. ينسق النظام نماذج اللغة الكبيرة من خلال توحيد التفاعلات عبر المزودين السحابيين والمحليين عبر واجهة موحدة. يدعم التوليد المعزز بالاسترجاع (RAG) من خلال دمج مصادر البيانات الخارجية وإضافات البحث في سير عمل النموذج. تشمل القدرات الإضافية إدارة الجلسات ذات الحالة لتتبع تاريخ المحادثة وبنية تعتمد على الإضافات لتوسيع أدوات الوكيل.
Standardizes requests and responses across different cloud and local language model providers using a single API layer.
AIOS is an LLM agent operating system and orchestration kernel designed to manage memory, resource scheduling, and tool execution for multiple autonomous AI agents. It serves as a comprehensive framework for developing and deploying agents, featuring a dedicated resource manager that coordinates model backends, GPU memory, and isolated kernel instances. The system distinguishes itself through a semantic memory engine that uses vector search and autonomous clustering for long-term knowledge management, and a semantic file system that allows users to control computer files and system operations
Provides a unified API that wraps multiple cloud APIs and local model weights for flexible backend switching.
Plano is an AI agent orchestrator and LLM gateway proxy that unifies access to multiple AI providers through a single interoperable interface. It functions as a model routing engine that decouples applications from specific vendors using semantic aliases, allowing traffic to be shifted between providers without modifying application code. The system distinguishes itself with intent-based agent routing, which directs prompts to specialized agents based on semantic analysis. It features an interceptor-based filter chain system that acts as guardrail middleware to enforce safety policies, rewrit
Provides a standardized API that offers a consistent execution interface for processing and streaming across different model providers.
Promptify عبارة عن مجموعة من الأدوات المصممة لتقييم النماذج، وإدارة المطالبات (prompts)، وتتبع تكلفة الرموز (tokens)، والاستخراج المهيكل، والوصول الموحد لبوابة API. يوفر واجهة موحدة لإدارة الطلبات والاستجابات عبر العديد من مزودي النماذج اللغوية الكبيرة. يتميز المشروع بمنصة إدارة مطالبات لهندسة وإصدار المطالبات مع التحقق من صحة المخرجات المهيكلة. يتضمن إطار عمل تقييم مخصص لقياس أداء النموذج باستخدام درجات الدقة والاستدعاء و f1 مقابل مجموعات البيانات المصنفة، إلى جانب متتبع تكلفة الرموز لمراقبة النفقات المالية لطلبات النموذج. تغطي المكتبة قدرات واسعة لمعالجة اللغة الطبيعية، بما في ذلك استخراج الكيانات المسماة، وتصنيف النصوص، والإجابة على الأسئلة. يدعم سير العمل عالي الحجم من خلال المعالجة المجمعة غير المتزامنة ويضمن اتساق البيانات عن طريق تحويل النص غير المهيكل إلى هياكل بيانات مكتوبة عبر التحقق من المخطط (schema validation).
Offers a unified abstraction layer to standardize requests and responses across different LLM providers.
Aigcpanel is a visual workflow automation tool and model lifecycle manager designed for generative AI media pipelines. It provides a unified interface to install, launch, and configure both local and remote AI model endpoints, acting as an orchestration platform for large language models and AI tools. The system features a drag-and-drop node editor for chaining AI models and scripts into automated processing pipelines. It distinguishes itself with a breakpoint-aware execution model that allows users to pause and resume long media tasks from specific points in the workflow. Additionally, it in
Provides a standardized API to ensure a consistent execution interface across different AI model providers.
LazyLLM is a multi-agent framework and orchestration engine designed for building complex AI applications. It provides a system for chaining large language models into sequential or parallel pipelines, utilizing a tool registry to convert standard functions into discoverable tools that models can invoke via reasoning. The project features an application deployment kit that enables hosting model workflows as web services with integrated chat interfaces and API gateways. It includes an infrastructure abstraction layer that allows users to switch between bare-metal servers, clusters, and public
Provides a standardized abstraction layer that maps diverse LLM provider APIs and local models to a consistent signature.
Koog is an LLM agent framework used to build autonomous entities that execute tool-based workflows. It utilizes a graph-based workflow engine to define agent behaviors and decision paths as a directed graph of nodes and edges. The framework distinguishes itself through a model provider orchestrator that enables dynamic switching, load balancing, and automatic fallbacks between different AI backends. It implements the Model Context Protocol to connect agents to remote tool servers and features a RAG memory system using vector embeddings to maintain long-term conversation context. The project
Provides a standardized execution interface that abstracts different cloud-based and local language model providers.
هذا المشروع عبارة عن إطار عمل شامل لبناء وتقييم وربط أنظمة الوكلاء المستقلين. يوفر مكتبة من الأنماط المعمارية الموحدة لتنفيذ تدفقات عمل الوكيل المعقدة، بما في ذلك تنسيق الوكلاء المتعددين، والتفكير التكراري، وإدارة الذاكرة. من خلال تقديم واجهة موحدة لموفري النماذج، يسمح إطار العمل بتنفيذ وكيل متسق عبر خدمات ذكاء اصطناعي مختلفة. يتميز إطار العمل بتركيزه على قياس الأداء الصارم والتحكم الحتمي. وهو يتضمن مجموعة من الأدوات لتقييم أداء الوكيل مقابل المهام الموحدة ومقاييس الجودة، مما يتيح مقارنة أنماط التصميم المختلفة. لضمان الموثوقية، يدمج النظام بوابات توجيه حتمية وحلقات تصحيح ذاتي تتحقق من إجراءات الوكيل وتصقل المخرجات مقابل معايير الجودة قبل التنفيذ الخارجي. تدعم البنية مجموعة واسعة من القدرات، بما في ذلك تكامل الأدوات لإنجاز المهام في العالم الحقيقي، والتوليد المعزز بالاسترجاع (RAG) للاستجابات المدركة للسياق، وإدارة الذاكرة المعيارية للحفاظ على المعلومات عبر الجلسات. يتم ربط هذه المكونات من خلال عقد تنفيذ موحد يضمن سلوكاً متسقاً بغض النظر عن النموذج الأساسي أو التكوين المعماري المحدد. يتم تنظيم المستودع كمجموعة من دفاتر Jupyter التي توضح هذه الأنماط ومنهجيات قياس الأداء.
Provides a standardized factory function to connect to various large language model providers, simplifying how applications request and receive data.