awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectAboutHow we rankPressMCP server
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
evaliphy avatar

evaliphy/evaliphy

0
View on GitHub↗
17 stars·9 forks·TypeScript·MIT·2 viewsevaliphy.com↗

Evaliphy

AI Evaluation Framework

Features

  • AI and LLM Testing - End-to-end testing approach for AI systems with HTML reporting.

Star history

Star history chart for evaliphy/evaliphyStar history chart for evaliphy/evaliphy

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Evaliphy

Similar open-source projects, ranked by how many features they share with Evaliphy.
  • mock-server/mockservermock-server avatar

    mock-server/mockserver

    4,900View on GitHub↗

    Mockserver is a multi-protocol mock server and API verification proxy used to simulate HTTP, gRPC, and WebSocket endpoints. It functions as a tool for testing client applications without relying on live backend services, providing a system to simulate chat completions and streaming responses for large language model integrations. The project automates behavior by generating request expectations and response behaviors from OpenAPI and Swagger specification files. It differentiates itself through a network traffic simulator that introduces latency, dropped connections, and failures to verify ho

    Java
    View on GitHub↗4,900
  • jwekavanagh/agentskepticjwekavanagh avatar

    jwekavanagh/agentskeptic

    0View on GitHub↗

    Tool effects vs read-only store facts.

    TypeScript
    View on GitHub↗0
  • promptfoo/promptfoopromptfoo avatar

    promptfoo/promptfoo

    10,529View on GitHub↗

    Promptfoo is an evaluation framework designed for testing, benchmarking, and red-teaming language models and agentic workflows. It provides a unified environment to run prompts against multiple providers, allowing developers to systematically validate model outputs against objective assertions, semantic similarity metrics, and custom grading rubrics. The platform distinguishes itself through a provider-agnostic execution layer and a stateful orchestrator capable of simulating multi-turn conversations and complex tool-use trajectories. It includes a dedicated adversarial mutation pipeline that

    TypeScriptcici-cdcicd
    View on GitHub↗10,529
  • tenro-ai/tenro-pythontenro-ai avatar

    tenro-ai/tenro-python

    6View on GitHub↗

    Simulate agent workflows and verify behavior without burning tokens.

    Python
    View on GitHub↗6
See all 5 alternatives to Evaliphy→

Frequently asked questions

What does evaliphy/evaliphy do?

AI Evaluation Framework

What are the main features of evaliphy/evaliphy?

The main features of evaliphy/evaliphy are: AI and LLM Testing.

What are some open-source alternatives to evaliphy/evaliphy?

Open-source alternatives to evaliphy/evaliphy include: mock-server/mockserver — Mockserver is a multi-protocol mock server and API verification proxy used to simulate HTTP, gRPC, and WebSocket… jwekavanagh/agentskeptic — Tool effects vs read-only store facts. promptfoo/promptfoo — Promptfoo is an evaluation framework designed for testing, benchmarking, and red-teaming language models and agentic… tenro-ai/tenro-python — Simulate agent workflows and verify behavior without burning tokens. voicetestdev/voicetest — Test harness for voice agents. Import from Retell, VAPI, Bland, LiveKit. Run autonomous simulations. Evaluate with LLM…