awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
raga-ai-hub avatar

raga-ai-hub/RagaAI-Catalyst

0
View on GitHub↗
16,150 نجوم·3,591 تفرعات·Python·Apache-2.0·11 مشاهداتcatalyst.raga.ai↗

RagaAI Catalyst

RagaAI-Catalyst is a suite of software implementation tools providing an SDK, dashboard, and platform for monitoring, debugging, red-teaming, and evaluating agentic AI workflows. It serves as an observability framework for tracing the execution paths of large language models and multi-agent systems.

The project distinguishes itself through a security suite for automated red-teaming and vulnerability scanning to detect biases, alongside a centralized prompt registry that decouples templates from application code. It further provides an evaluation platform that combines synthetic data generation with custom metric frameworks to quantify model accuracy and reliability.

The system covers broad operational domains including agent behavioral observability, prompt lifecycle management, and the application of output guardrails to block undesirable content. Its monitoring capabilities include trace-based execution graphing, timeline-based event sequencing, and diagnostic tools for analyzing multi-agent interaction flows.

The core functionality is delivered via a Python library for recording tool calls and decision-making processes.

Features

  • Agent Execution Tracing - Implements a system for tracing agent reasoning, tool calls, and decision-making processes to debug complex workflows.
  • Agent Observability - Provides a comprehensive framework for monitoring and tracing the decision-making processes of autonomous agents.
  • Agent Debugging Tools - Provides specialized tools for analyzing interaction timelines and execution graphs to debug agent logic.
  • Agent Tracing SDKs - Provides a Python library for recording tool calls and decision-making processes to debug agent behaviors.
  • LLM Observability - Serves as a complete framework for tracing and monitoring execution paths of LLMs and multi-agent workflows.
  • Model Evaluation Metrics - Provides frameworks for quantifying model accuracy and reliability using specific performance benchmarks.
  • Model Red-Teaming - Provides a security suite for automated red-teaming and vulnerability scanning to detect model biases.
  • Output Guardrails - Implements a validation layer to intercept and block harmful or undesirable model responses.
  • Prompt Management - Organizes and maintains a central library of prompts to ensure consistency across environments.
  • Prompt Registries - Provides a centralized system for defining and versioning prompt templates independently of application code.
  • AI Red Teaming - Scans models for vulnerabilities and biases using synthetic test cases and automated detectors.
  • Adversarial Red Teaming Toolkits - Runs automated adversarial test cases to detect biases and safety failures in AI models.
  • LLM Evaluation - Measures the accuracy and reliability of LLM outputs using specialized metrics and automated judges.
  • Synthetic Data Generators - Produces artificial datasets used to facilitate the testing and evaluation of AI applications.
  • Agentic Workflow Generators - Generates artificial datasets used for stress testing and evaluating complex agentic AI workflows.
  • Model Output Safeguarding - Implements output guardrails to block undesirable content and ensure model responses align with safety guidelines.
  • Trace-Based Flow Visualizers - Captures sequential tool calls and model interactions to visualize the logic flow as a directed graph.
  • Agent Interaction Timelines - Orders asynchronous agent interactions on a linear timeline to identify latency bottlenecks and race conditions.
  • Agent Interaction Dashboards - Provides a visual interface for monitoring and analyzing execution graphs in multi-agent systems.
  • Model Evaluation and Benchmarking - Platform for managing and optimizing LLM project performance.

سجل النجوم

مخطط تاريخ النجوم لـ raga-ai-hub/ragaai-catalystمخطط تاريخ النجوم لـ raga-ai-hub/ragaai-catalyst

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Start searching with AI

بدائل مفتوحة المصدر لـ RagaAI Catalyst

مشاريع مفتوحة المصدر مشابهة، مرتبة حسب عدد الميزات المشتركة مع RagaAI Catalyst.
  • confident-ai/deepevalالصورة الرمزية لـ confident-ai

    confident-ai/deepeval

    13,733عرض على GitHub↗

    Deepeval is a framework for testing and evaluating large language model applications. It provides a suite of tools for executing automated regression tests, validating model output quality against defined standards, and tracing the execution of complex agent workflows. By integrating these capabilities into development pipelines, the platform ensures consistent performance and reliability throughout the software lifecycle. The platform distinguishes itself through its focus on programmatic validation and observability. It utilizes secondary language models to score output quality and employs

    Pythonevaluation-frameworkevaluation-metricsllm-evaluation
    عرض على GitHub↗13,733
  • agenta-ai/agentaالصورة الرمزية لـ Agenta-AI

    Agenta-AI/agenta

    3,860عرض على GitHub↗

    Agenta is a Prompt Ops lifecycle manager and prompt management platform that decouples prompt engineering from application code. It serves as a centralized system for developing, versioning, and deploying prompt templates and model configurations across different environments. The platform functions as an AI agent orchestrator with a visual interface for building agent workflows and connecting models to external tools. It further acts as an evaluation framework and observability tool, utilizing OpenTelemetry to capture execution traces, monitor latency, and track token costs. The system cove

    TypeScriptagentsevaluationllm-as-a-judge
    عرض على GitHub↗3,860
  • comet-ml/opikالصورة الرمزية لـ comet-ml

    comet-ml/opik

    17,787عرض على GitHub↗

    Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It provides a centralized environment for tracing execution flows, managing prompt templates, and monitoring production performance, allowing teams to gain visibility into complex model interactions and tool usage without requiring manual application code changes. The platform distinguishes itself through its integrated approach to the AI development lifecycle, combining distributed trace instrumentation with automated evaluation frameworks. It supports model-as-a-judge scoring, syn

    Pythonevaluationhacktoberfesthacktoberfest2025
    عرض على GitHub↗17,787
  • promptfoo/promptfooالصورة الرمزية لـ promptfoo

    promptfoo/promptfoo

    10,529عرض على GitHub↗

    Promptfoo is an evaluation framework designed for testing, benchmarking, and red-teaming language models and agentic workflows. It provides a unified environment to run prompts against multiple providers, allowing developers to systematically validate model outputs against objective assertions, semantic similarity metrics, and custom grading rubrics. The platform distinguishes itself through a provider-agnostic execution layer and a stateful orchestrator capable of simulating multi-turn conversations and complex tool-use trajectories. It includes a dedicated adversarial mutation pipeline that

    TypeScriptcici-cdcicd
    عرض على GitHub↗10,529
عرض جميع البدائل الـ 30 لـ RagaAI Catalyst→

الأسئلة الشائعة

ما هي وظيفة raga-ai-hub/ragaai-catalyst؟

RagaAI-Catalyst is a suite of software implementation tools providing an SDK, dashboard, and platform for monitoring, debugging, red-teaming, and evaluating agentic AI workflows. It serves as an observability framework for tracing the execution paths of large language models and multi-agent systems.

ما هي الميزات الرئيسية لـ raga-ai-hub/ragaai-catalyst؟

الميزات الرئيسية لـ raga-ai-hub/ragaai-catalyst هي: Agent Execution Tracing, Agent Observability, Agent Debugging Tools, Agent Tracing SDKs, LLM Observability, Model Evaluation Metrics, Model Red-Teaming, Output Guardrails.

ما هي البدائل مفتوحة المصدر لـ raga-ai-hub/ragaai-catalyst؟

تشمل البدائل مفتوحة المصدر لـ raga-ai-hub/ragaai-catalyst: confident-ai/deepeval — Deepeval is a framework for testing and evaluating large language model applications. It provides a suite of tools for… agenta-ai/agenta — Agenta is a Prompt Ops lifecycle manager and prompt management platform that decouples prompt engineering from… comet-ml/opik — Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It… promptfoo/promptfoo — Promptfoo is an evaluation framework designed for testing, benchmarking, and red-teaming language models and agentic… evidentlyai/evidently — Evidently is an AI observability platform and evaluation framework designed to quantify the performance of machine… arize-ai/phoenix — Arize Phoenix is an LLM observability platform and evaluation framework designed to capture execution traces and…