1 Repo
Comprehensive toolkits for training and deploying audio models based on latent diffusion.
Distinct from Latent Diffusion Models: Distinct from Latent Diffusion Models: refers to the full framework for training and deployment, not just the model architecture.
Explore 1 awesome GitHub repository matching artificial intelligence & ml · Audio Latent Diffusion Frameworks. Refine with filters or upvote what's useful.
Stable-audio-tools is a toolkit for training and deploying latent diffusion models for high-fidelity audio synthesis. It provides a framework for generating audio by iteratively refining noise within a compressed latent space, using specialized encoders to preserve temporal and spectral features of the audio signal. The project features a system for adapting pre-trained audio checkpoints to new datasets through modular initialization and configuration files. It includes utilities for weight extraction and inference model export, which remove training metadata and optimizer states to create li
Provides a complete toolkit for training and deploying generative audio models using latent diffusion and compressed spaces.