awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
aliutkus avatar

aliutkus/speechmetrics

0
View on GitHub↗
1,051 stars·174 forks·Python·MIT·11 views

Speechmetrics

A wrapper around speech quality metrics MOSNet, BSSEval, STOI, PESQ, SRMR, SISDR

Features

  • Speech Processing - Evaluation metrics for speech quality assessment.

Star history

Star history chart for aliutkus/speechmetricsStar history chart for aliutkus/speechmetrics

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Speechmetrics

Similar open-source projects, ranked by how many features they share with Speechmetrics.
  • aishell-foundation/dacidianaishell-foundation avatar

    aishell-foundation/DaCiDian

    301View on GitHub↗

    DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR)

    Python
    View on GitHub↗301
  • as-ideas/transformerttsas-ideas avatar

    as-ideas/TransformerTTS

    1,161View on GitHub↗

    🤖💬 Transformer TTS: Implementation of a non-autoregressive Transformer based neural network for text to speech.

    Python
    View on GitHub↗1,161
  • boson-ai/higgs-audioboson-ai avatar

    boson-ai/higgs-audio

    7,919View on GitHub↗

    Higgs-audio is a generative text-to-speech engine that transforms text into natural conversational speech using large language model architectures. It functions as a multilingual speech synthesizer capable of generating high-fidelity audio across different languages with control over emotional tone and prosody. The system includes a voice cloning tool that creates synthetic replicas of specific speakers from short audio samples without requiring extensive model training. It also provides a streaming audio API designed to deliver generated speech incrementally to minimize playback delay. The

    Python
    View on GitHub↗7,919
  • 2noise/chattts2noise avatar

    2noise/ChatTTS

    39,464View on GitHub↗

    ChatTTS is a conversational text-to-speech generative model designed to convert written dialogue into natural sounding audio. It functions as a multilingual speech synthesis framework capable of producing human-like audio across different languages and speaker profiles. The system is distinguished by its ability to generate interactive dialogue with realistic vocal nuances. It utilizes a speech nuance controller to insert specific tokens that trigger non-verbal elements, such as laughter, pauses, and interjections, during the synthesis process. The project includes a streaming audio generato

    Pythonagentchatchatgpt
    View on GitHub↗39,464
See all 30 alternatives to Speechmetrics→

Frequently asked questions

What does aliutkus/speechmetrics do?

A wrapper around speech quality metrics MOSNet, BSSEval, STOI, PESQ, SRMR, SISDR

What are the main features of aliutkus/speechmetrics?

The main features of aliutkus/speechmetrics are: Speech Processing.

What are some open-source alternatives to aliutkus/speechmetrics?

Open-source alternatives to aliutkus/speechmetrics include: aishell-foundation/dacidian — DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR). as-ideas/transformertts — 🤖💬 Transformer TTS: Implementation of a non-autoregressive Transformer based neural network for text to speech. boson-ai/higgs-audio — Higgs-audio is a generative text-to-speech engine that transforms text into natural conversational speech using large… bytedance/megatts3 — MegaTTS3 is a bilingual speech synthesis system that generates natural-sounding speech in Chinese and English,… emotional-text-to-speech/dl-for-emo-tts — :computer: :robot: A summary on our attempts at using Deep Learning approaches for Emotional Text to Speech :speaker:. 2noise/chattts — ChatTTS is a conversational text-to-speech generative model designed to convert written dialogue into natural sounding…