awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
archinetai avatar

archinetai/audio-diffusion-pytorch

0
View on GitHub↗
2,100 stele·177 fork-uri·Python·MIT·7 vizualizări

Audio Diffusion Pytorch

Audio generation using diffusion models, in PyTorch.

Features

  • Audio Generation - Audio generation using diffusion models in PyTorch.

Istoric stele

Graficul istoricului de stele pentru archinetai/audio-diffusion-pytorchGraficul istoricului de stele pentru archinetai/audio-diffusion-pytorch

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Întrebări frecvente

Ce face archinetai/audio-diffusion-pytorch?

Audio generation using diffusion models, in PyTorch.

Care sunt principalele funcționalități ale archinetai/audio-diffusion-pytorch?

Principalele funcționalități ale archinetai/audio-diffusion-pytorch sunt: Audio Generation.

Care sunt câteva alternative open-source pentru archinetai/audio-diffusion-pytorch?

Alternativele open-source pentru archinetai/audio-diffusion-pytorch includ: open-mmlab/amphion — Amphion is an audio generation toolkit designed for the research and development of models that synthesize speech,… kittenml/kittentts — KittenTTS is a neural text-to-speech engine and text-to-audio synthesis tool that converts written text into spoken… rsxdalv/tts-webui — TTS-WebUI is a web interface and speech synthesis manager designed to convert written text into spoken audio files. It… aigc-audio/audiogpt — AudioGPT is an LLM-driven audio framework and processing suite that uses large language models to orchestrate neural… magenta/magenta — Magenta is a comprehensive toolkit for training, synthesizing, and performing music through neural models and… mubertai/mubert-text-to-music — A simple notebook demonstrating prompt-based music generation via Mubert API.

Alternative open-source pentru Audio Diffusion Pytorch

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu Audio Diffusion Pytorch.
  • open-mmlab/amphionAvatar open-mmlab

    open-mmlab/Amphion

    9,844Vezi pe GitHub↗

    Amphion is an audio generation toolkit designed for the research and development of models that synthesize speech, music, and environmental sound effects. It provides a standardized framework for reproducible audio synthesis, incorporating a text-to-speech engine and a voice conversion framework. The project specializes in transforming audio identities, allowing for the modification of speaker accents and voice identities while preserving original rhythm and style. It also includes capabilities for singing voice synthesis and the generation of environmental soundscapes from text descriptions

    Pythonaudio-generationaudio-synthesisaudioldm
    Vezi pe GitHub↗9,844
  • rsxdalv/tts-webuiAvatar rsxdalv

    rsxdalv/TTS-WebUI

    2,980Vezi pe GitHub↗

    TTS-WebUI is a web interface and speech synthesis manager designed to convert written text into spoken audio files. It serves as a self-hosted audio AI suite that allows users to configure speech synthesis models, manage speaker profiles, and generate audio through a graphical dashboard. The system functions as both a visual manager and a generative audio API, providing standardized endpoints and OpenAI-compatible request formats for external applications to trigger synthesis programmatically. It includes a plugin-based extension system that allows new tools and models to be added via externa

    TypeScriptace-stepaiaudio-generation
    Vezi pe GitHub↗2,980
  • kittenml/kittenttsAvatar KittenML

    KittenML/KittenTTS

    10,044Vezi pe GitHub↗

    KittenTTS is a neural text-to-speech engine and text-to-audio synthesis tool that converts written text into spoken audio using lightweight neural network models. It functions as both a speech synthesizer and an audio file generator, producing spoken audio for offline playback. The system includes a text normalization processor that expands numbers and abbreviations into full spoken words to improve the naturalness of the synthesized speech. It supports diverse voice options and provides the ability to adjust playback speed.

    Python
    Vezi pe GitHub↗10,044
  • facebookresearch/audiocraftAvatar facebookresearch

    facebookresearch/audiocraft

    23,379Vezi pe GitHub↗

    Audiocraft is a deep learning audio library and machine learning framework designed for training, fine-tuning, and evaluating generative models for music and sound effects. It functions as a text-to-music generative model and a neural audio codec, providing the tools necessary to compress audio signals into discrete representations and synthesize high-fidelity waveforms from textual descriptions. The framework is distinguished by its ability to combine multiple conditioning signals, allowing for the generation of audio based on text prompts, melodic excerpts, or style-based audio clips. It al

    Jupyter Notebook
    Vezi pe GitHub↗23,379
Vezi toate cele 9 alternative pentru Audio Diffusion Pytorch→