awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
as-ideas avatar

as-ideas/TransformerTTSArchived

0
View on GitHub↗
1,161 stars·222 forks·Python·4 viewsas-ideas.github.io/TransformerTTS↗

TransformerTTS

🤖💬 Transformer TTS: Implementation of a non-autoregressive Transformer based neural network for text to speech.

Features

  • Speech Processing - Transformer-based text-to-speech implementation.

Star history

Star history chart for as-ideas/transformerttsStar history chart for as-ideas/transformertts

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to TransformerTTS

Similar open-source projects, ranked by how many features they share with TransformerTTS.
  • aishell-foundation/dacidianaishell-foundation avatar

    aishell-foundation/DaCiDian

    301View on GitHub↗

    DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR)

    Python
    View on GitHub↗301
  • aliutkus/speechmetricsaliutkus avatar

    aliutkus/speechmetrics

    1,051View on GitHub↗

    A wrapper around speech quality metrics MOSNet, BSSEval, STOI, PESQ, SRMR, SISDR

    Python
    View on GitHub↗1,051
  • boson-ai/higgs-audioboson-ai avatar

    boson-ai/higgs-audio

    7,919View on GitHub↗

    Higgs-audio is a generative text-to-speech engine that transforms text into natural conversational speech using large language model architectures. It functions as a multilingual speech synthesizer capable of generating high-fidelity audio across different languages with control over emotional tone and prosody. The system includes a voice cloning tool that creates synthetic replicas of specific speakers from short audio samples without requiring extensive model training. It also provides a streaming audio API designed to deliver generated speech incrementally to minimize playback delay. The

    Python
    View on GitHub↗7,919
  • 2noise/chattts2noise avatar

    2noise/ChatTTS

    39,464View on GitHub↗

    ChatTTS is a conversational text-to-speech generative model designed to convert written dialogue into natural sounding audio. It functions as a multilingual speech synthesis framework capable of producing human-like audio across different languages and speaker profiles. The system is distinguished by its ability to generate interactive dialogue with realistic vocal nuances. It utilizes a speech nuance controller to insert specific tokens that trigger non-verbal elements, such as laughter, pauses, and interjections, during the synthesis process. The project includes a streaming audio generato

    Pythonagentchatchatgpt
    View on GitHub↗39,464
See all 30 alternatives to TransformerTTS→

Frequently asked questions

What does as-ideas/transformertts do?

🤖💬 Transformer TTS: Implementation of a non-autoregressive Transformer based neural network for text to speech.

What are the main features of as-ideas/transformertts?

The main features of as-ideas/transformertts are: Speech Processing.

What are some open-source alternatives to as-ideas/transformertts?

Open-source alternatives to as-ideas/transformertts include: aishell-foundation/dacidian — DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR). aliutkus/speechmetrics — A wrapper around speech quality metrics MOSNet, BSSEval, STOI, PESQ, SRMR, SISDR. boson-ai/higgs-audio — Higgs-audio is a generative text-to-speech engine that transforms text into natural conversational speech using large… bytedance/megatts3 — MegaTTS3 is a bilingual speech synthesis system that generates natural-sounding speech in Chinese and English,… emotional-text-to-speech/dl-for-emo-tts — :computer: :robot: A summary on our attempts at using Deep Learning approaches for Emotional Text to Speech :speaker:. 2noise/chattts — ChatTTS is a conversational text-to-speech generative model designed to convert written dialogue into natural sounding…