awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 个仓库

Awesome GitHub RepositoriesAudio Performance Benchmarks

Standardized tests and metrics for evaluating the quality and performance of speech and audio models.

Distinct from Model Performance Benchmarking: Specifically targets audio and speech fidelity and accuracy, whereas the parent is general model performance benchmarking.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Audio Performance Benchmarks. Refine with filters or upvote what's useful.

Awesome Audio Performance Benchmarks GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • moonshotai/kimi-audioMoonshotAI 的头像

    MoonshotAI/Kimi-Audio

    4,492在 GitHub 上查看↗

    Kimi-Audio is a large language model audio foundation model designed to understand audio input and generate high-fidelity speech responses in real time. It functions as a unified system encompassing a text-to-speech synthesis engine and a speech-to-text transcription tool. The project enables real-time audio conversations through a multi-modal conversation loop and chunk-wise streaming detokenization to reduce playback latency. It provides controls over speech speed, accent, and emotional tone during conversational audio generation. The system covers audio intelligence capabilities, includin

    Ships a benchmarking harness with standardized metrics and side-by-side inference recipes for audio models.

    Python
    在 GitHub 上查看↗4,492
  • evolvinglmms-lab/lmms-evalEvolvingLMMs-Lab 的头像

    EvolvingLMMs-Lab/lmms-eval

    3,701在 GitHub 上查看↗

    lmms-eval is a benchmarking system and performance analysis suite designed to measure the capabilities of large multimodal models. It provides a framework for evaluating models across text, image, audio, and video datasets, serving as a multimodal dataset orchestrator and benchmarking tool to quantify accuracy and efficiency. The project distinguishes itself through a unified multimodal message protocol that structures diverse media inputs for consistent model consumption. It features specialized benchmarking for audio, video, visual, document, and spatial reasoning, alongside tools for model

    Assesses model capabilities in speech recognition, speech translation, and audio-based question answering.

    Pythonagiaudio-evaluationbenchmark
    在 GitHub 上查看↗3,701
  1. Home
  2. Artificial Intelligence & ML
  3. Machine Learning
  4. Infrastructure
  5. Model Evaluation and Analysis
  6. Model Analysis
  7. Model Performance Benchmarking
  8. Audio Performance Benchmarks