awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
future-agi avatar

future-agi/future-agi

0
View on GitHub↗
1,175 stars·259 forks·Python·Apache-2.0·5 viewsfutureagi.com↗

Future Agi

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

Features

  • Agent Frameworks - End-to-end platform for agent engineering and optimization.
  • AI Observability Tools - End-to-end platform for LLM tracing, evaluations, and agent simulations.
  • Model Evaluation and Benchmarking - End-to-end platform for agent engineering, tracing, and guardrails.
  • Monitoring and Observability - Self-hostable LLMOps platform for tracing and evaluation.

Star history

Star history chart for future-agi/future-agiStar history chart for future-agi/future-agi

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Future Agi

Similar open-source projects, ranked by how many features they share with Future Agi.
  • langfuse/langfuselangfuse avatar

    langfuse/langfuse

    29,190View on GitHub↗

    Langfuse is an open-source observability and evaluation platform designed for language model applications. It provides a centralized system for tracking execution traces, monitoring performance metrics, and managing prompt templates. By capturing hierarchical units of work and telemetry data, the platform enables developers to debug complex application lifecycles and analyze token usage, latency, and model interactions in production environments. The platform distinguishes itself through an integrated evaluation framework that allows for systematic benchmarking and automated scoring of model

    TypeScriptanalyticsautogenevaluation
    View on GitHub↗29,190
  • helicone/heliconeHelicone avatar

    Helicone/helicone

    5,830View on GitHub↗

    Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with large language models. By acting as a reverse-proxy, it provides a centralized layer for routing requests across multiple AI providers, allowing developers to maintain consistent application logic while gaining deep visibility into model performance, usage, and costs. The platform distinguishes itself through a robust suite of traffic management and prompt engineering tools. It enables policy-driven control, including automatic failover between providers, rate limiting, and edge-b

    TypeScript
    View on GitHub↗5,830
  • comet-ml/opikcomet-ml avatar

    comet-ml/opik

    17,787View on GitHub↗

    Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It provides a centralized environment for tracing execution flows, managing prompt templates, and monitoring production performance, allowing teams to gain visibility into complex model interactions and tool usage without requiring manual application code changes. The platform distinguishes itself through its integrated approach to the AI development lifecycle, combining distributed trace instrumentation with automated evaluation frameworks. It supports model-as-a-judge scoring, syn

    Pythonevaluationhacktoberfesthacktoberfest2025
    View on GitHub↗17,787
  • lmnr-ai/lmnrlmnr-ai avatar

    lmnr-ai/lmnr

    2,608View on GitHub↗

    Lmnr is an LLM observability platform and evaluation framework designed for tracing, logging, and monitoring language model executions. It provides the tools necessary to debug agent behavior, analyze performance, and identify failure patterns in AI agents. The platform differentiates itself through a trace-to-dataset pipeline that converts production logs into labeled test sets for regression testing. It includes a prompt-variant replay engine to compare different prompts or models side-by-side and a state-cached debugging system to replay agent loops without restarting the process. The sys

    TypeScriptagentsaiai-observability
    View on GitHub↗2,608
See all 30 alternatives to Future Agi→

Frequently asked questions

What does future-agi/future-agi do?

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

What are the main features of future-agi/future-agi?

The main features of future-agi/future-agi are: Agent Frameworks, AI Observability Tools, Model Evaluation and Benchmarking, Monitoring and Observability.

What are some open-source alternatives to future-agi/future-agi?

Open-source alternatives to future-agi/future-agi include: lmnr-ai/lmnr — Lmnr is an LLM observability platform and evaluation framework designed for tracing, logging, and monitoring language… langfuse/langfuse — Langfuse is an open-source observability and evaluation platform designed for language model applications. It provides… comet-ml/opik — Opik is an observability and evaluation platform designed for generative AI applications and agentic workflows. It… helicone/helicone — Helicone is an AI gateway and observability platform designed to intercept, manage, and monitor interactions with… facebookresearch/parlai — ParlAI is a conversational AI research framework designed for training, evaluating, and sharing dialogue models using… infrasys-ai/aiinfra.