awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetÀ proposNotre méthodologiePresseServeur MCP
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
langwatch avatar

langwatch/langwatch

0
View on GitHub↗
3,307 stars·323 forks·TypeScript·Apache-2.0·5 vueslangwatch.ai↗

Langwatch

The platform for LLM evaluations and AI agent testing

Features

  • Application Development - Platform for LLM observability and prompt optimization.
  • Application Services - Observability and evaluation tool for LLM apps.
  • Generative AI - Listed in the “Generative AI” section of the Free For Dev awesome list.
  • Inference Optimization - Studio for evaluating and optimizing LLM workflows.
  • LLM Observability and Evaluation - Platform for monitoring, analytics, and prompt optimization.
  • Model Evaluation and Benchmarking - Platform for monitoring, experimenting, and improving LLM pipelines.
  • Model Visualization - Visualizes LLM evaluation experiments and pipeline optimizations.
  • Observability and Evaluation - Platform for monitoring and optimizing LLM application performance.
  • Low Code Interfaces - Platform for monitoring and optimizing LLM performance with visual tools.

Historique des stars

Graphique de l'historique des stars pour langwatch/langwatchGraphique de l'historique des stars pour langwatch/langwatch

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Alternatives open source à Langwatch

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec Langwatch.
  • langfuse/langfuseAvatar de langfuse

    langfuse/langfuse

    29,190Voir sur GitHub↗

    Langfuse is an open-source observability and evaluation platform designed for language model applications. It provides a centralized system for tracking execution traces, monitoring performance metrics, and managing prompt templates. By capturing hierarchical units of work and telemetry data, the platform enables developers to debug complex application lifecycles and analyze token usage, latency, and model interactions in production environments. The platform distinguishes itself through an integrated evaluation framework that allows for systematic benchmarking and automated scoring of model

    TypeScriptanalyticsautogenevaluation
    Voir sur GitHub↗29,190
  • comet-ml/opikAvatar de comet-ml

    comet-ml/opik

    17,787Voir sur GitHub↗

    Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It provides a centralized environment for tracing execution flows, managing prompt templates, and monitoring production performance, allowing teams to gain visibility into complex model interactions and tool usage without requiring manual application code changes. The platform distinguishes itself through its integrated approach to the AI development lifecycle, combining distributed trace instrumentation with automated evaluation frameworks. It supports model-as-a-judge scoring, syn

    Pythonevaluationhacktoberfesthacktoberfest2025
    Voir sur GitHub↗17,787
  • arize-ai/phoenixAvatar de Arize-ai

    Arize-ai/phoenix

    8,605Voir sur GitHub↗

    Arize Phoenix is an LLM observability platform and evaluation framework designed to capture execution traces and monitor large language model applications. It serves as a prompt management system for versioning and testing templates, and as a self-hosted AI operations infrastructure for managing telemetry and experiments. The platform differentiates itself through a specialized embedding visualization tool used to detect data drift and optimize vector search. It provides a comprehensive evaluation suite that utilizes judge-based evaluators and ground-truth datasets to score model outputs, and

    Jupyter Notebookagentsai-monitoringai-observability
    Voir sur GitHub↗8,605
  • evidentlyai/evidentlyAvatar de evidentlyai

    evidentlyai/evidently

    7,137Voir sur GitHub↗

    Evidently is an AI observability platform and evaluation framework designed to quantify the performance of machine learning models and large language models. It functions as a monitoring tool for detecting data drift and quality degradation in tabular datasets, while providing a specialized analyzer for the faithfulness and correctness of retrieval augmented generation systems. The project distinguishes itself through an evaluation framework that utilizes judge models and custom rubrics to score language model outputs. It includes tools for iterative prompt optimization and the generation of

    Jupyter Notebookdata-driftdata-qualitydata-science
    Voir sur GitHub↗7,137
Voir les 30 alternatives à Langwatch→

Questions fréquentes

Que fait langwatch/langwatch ?

The platform for LLM evaluations and AI agent testing

Quelles sont les fonctionnalités principales de langwatch/langwatch ?

Les fonctionnalités principales de langwatch/langwatch sont : Application Development, Application Services, Generative AI, Inference Optimization, LLM Observability and Evaluation, Model Evaluation and Benchmarking, Model Visualization, Observability and Evaluation.

Quelles sont les alternatives open-source à langwatch/langwatch ?

Les alternatives open-source à langwatch/langwatch incluent : comet-ml/opik — Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It… langfuse/langfuse — Langfuse is an open-source observability and evaluation platform designed for language model applications. It provides… helicone/helicone — Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with… evidentlyai/evidently — Evidently is an AI observability platform and evaluation framework designed to quantify the performance of machine… arize-ai/phoenix — Arize Phoenix is an LLM observability platform and evaluation framework designed to capture execution traces and… latitude-dev/latitude-llm — This project is a self-hosted AI monitoring stack that functions as an LLM observability platform, AI evaluation…