awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

10 repository-uri

Awesome GitHub RepositoriesAudio Generation

Tools and models for synthesizing music, sound effects, and speech.

Explore 10 awesome GitHub repositories matching part of an awesome list · Audio Generation. Refine with filters or upvote what's useful.

Awesome Audio Generation GitHub Repositories

Găsește cele mai bune repo-uri cu AI.Vom căuta cele mai potrivite repository-uri folosind AI.
  • facebookresearch/audiocraftAvatar facebookresearch

    facebookresearch/audiocraft

    23,379Vezi pe GitHub↗

    Audiocraft is a deep learning audio library and machine learning framework designed for training, fine-tuning, and evaluating generative models for music and sound effects. It functions as a text-to-music generative model and a neural audio codec, providing the tools necessary to compress audio signals into discrete representations and synthesize high-fidelity waveforms from textual descriptions. The framework is distinguished by its ability to combine multiple conditioning signals, allowing for the generation of audio based on text prompts, melodic excerpts, or style-based audio clips. It al

    Library for audio processing and generation using deep learning.

    Jupyter Notebook
    Vezi pe GitHub↗23,379
  • magenta/magentaAvatar magenta

    magenta/magenta

    19,778Vezi pe GitHub↗

    Magenta is a comprehensive toolkit for training, synthesizing, and performing music through neural models and hardware-integrated engines. It functions as a machine learning framework that enables the generation, manipulation, and real-time performance of audio, providing the structural foundations for musical intelligence through hierarchical sequence modeling and symbolic processing. The project distinguishes itself by enabling real-time, low-latency neural audio synthesis that can be integrated directly into professional digital audio workstations. It supports interactive musical jamming a

    Research project for music and art generation with machine intelligence.

    Python
    Vezi pe GitHub↗19,778
  • aigc-audio/audiogptAvatar AIGC-Audio

    AIGC-Audio/AudioGPT

    10,174Vezi pe GitHub↗

    AudioGPT is an LLM-driven audio framework and processing suite that uses large language models to orchestrate neural audio pipelines. It functions as a multimodal audio generator and processing system, integrating a collection of pretrained models to handle speech synthesis, sound generation, and audio manipulation. The system is distinguished by its ability to generate audio from diverse inputs, including text and images, and its capacity to produce synchronized talking head videos. It also operates as a neural speech translator, converting spoken language between different tongues while pre

    Framework for understanding and generating speech, music, and sound.

    Pythonaudiogptmusic
    Vezi pe GitHub↗10,174
  • kittenml/kittenttsAvatar KittenML

    KittenML/KittenTTS

    10,044Vezi pe GitHub↗

    KittenTTS is a neural text-to-speech engine and text-to-audio synthesis tool that converts written text into spoken audio using lightweight neural network models. It functions as both a speech synthesizer and an audio file generator, producing spoken audio for offline playback. The system includes a text normalization processor that expands numbers and abbreviations into full spoken words to improve the naturalness of the synthesized speech. It supports diverse voice options and provides the ability to adjust playback speed.

    Functions as a generator that writes synthesized speech directly to audio files.

    Python
    Vezi pe GitHub↗10,044
  • open-mmlab/amphionAvatar open-mmlab

    open-mmlab/Amphion

    9,844Vezi pe GitHub↗

    Amphion is an audio generation toolkit designed for the research and development of models that synthesize speech, music, and environmental sound effects. It provides a standardized framework for reproducible audio synthesis, incorporating a text-to-speech engine and a voice conversion framework. The project specializes in transforming audio identities, allowing for the modification of speaker accents and voice identities while preserving original rhythm and style. It also includes capabilities for singing voice synthesis and the generation of environmental soundscapes from text descriptions

    Provides a comprehensive toolkit for synthesizing speech, music, and environmental sound effects.

    Pythonaudio-generationaudio-synthesisaudioldm
    Vezi pe GitHub↗9,844
  • lucidrains/musiclm-pytorchAvatar lucidrains

    lucidrains/musiclm-pytorch

    3,290Vezi pe GitHub↗

    Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch

    Implementation of a state-of-the-art music generation model.

    Python
    Vezi pe GitHub↗3,290
  • rsxdalv/tts-webuiAvatar rsxdalv

    rsxdalv/TTS-WebUI

    2,980Vezi pe GitHub↗

    TTS-WebUI is a web interface and speech synthesis manager designed to convert written text into spoken audio files. It serves as a self-hosted audio AI suite that allows users to configure speech synthesis models, manage speaker profiles, and generate audio through a graphical dashboard. The system functions as both a visual manager and a generative audio API, providing standardized endpoints and OpenAI-compatible request formats for external applications to trigger synthesis programmatically. It includes a plugin-based extension system that allows new tools and models to be added via externa

    Offers tools for synthesizing music, sound effects, and speech within a unified web interface.

    TypeScriptace-stepaiaudio-generation
    Vezi pe GitHub↗2,980
  • mubertai/mubert-text-to-musicAvatar MubertAI

    MubertAI/Mubert-Text-to-Music

    2,731Vezi pe GitHub↗

    A simple notebook demonstrating prompt-based music generation via Mubert API

    Notebook demonstrating prompt-based music generation via API.

    Jupyter Notebook
    Vezi pe GitHub↗2,731
  • archinetai/audio-diffusion-pytorchAvatar archinetai

    archinetai/audio-diffusion-pytorch

    2,100Vezi pe GitHub↗

    Audio generation using diffusion models, in PyTorch.

    Audio generation using diffusion models in PyTorch.

    Python
    Vezi pe GitHub↗2,100
  • archinetai/audio-ai-timelineAvatar archinetai

    archinetai/audio-ai-timeline

    1,910Vezi pe GitHub↗

    A timeline of the latest AI models for audio generation, starting in 2023!

    Chronological record of recent audio generation model releases.

    artificial-intelligenceaudio-generationmachine-learning
    Vezi pe GitHub↗1,910
  1. Home
  2. Part of an Awesome List
  3. AI & Machine Learning
  4. Audio Generation