awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 repositorios

Awesome GitHub RepositoriesMultimodal Generation

Audio generation models that utilize non-audio inputs such as images or text to synthesize soundscapes.

Distinct from Audio Generation Models: Focuses on cross-modal input mapping specifically, whereas Audio Generation Models is the broader category for any audio synthesis.

Explore 2 awesome GitHub repositories matching artificial intelligence & ml · Multimodal Generation. Refine with filters or upvote what's useful.

Awesome Multimodal Generation GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • aigc-audio/audiogptAvatar de AIGC-Audio

    AIGC-Audio/AudioGPT

    10,174Ver en GitHub↗

    AudioGPT is an LLM-driven audio framework and processing suite that uses large language models to orchestrate neural audio pipelines. It functions as a multimodal audio generator and processing system, integrating a collection of pretrained models to handle speech synthesis, sound generation, and audio manipulation. The system is distinguished by its ability to generate audio from diverse inputs, including text and images, and its capacity to produce synchronized talking head videos. It also operates as a neural speech translator, converting spoken language between different tongues while pre

    Implements a system that generates soundscapes and audio clips from visual images or natural language descriptions.

    Pythonaudiogptmusic
    Ver en GitHub↗10,174
  • ml-explore/mlx-examplesAvatar de ml-explore

    ml-explore/mlx-examples

    8,254Ver en GitHub↗

    This repository provides a collection of reference implementations and code examples for training and deploying machine learning models using the MLX framework. It serves as a practical guide for executing distributed training, fine-tuning large language models, converting model weights, and implementing multimodal generative workflows. The project distinguishes itself through specialized examples for local hardware execution, featuring weight quantization to reduce memory usage and low-rank adaptation for parameter-efficient fine-tuning. It also includes scripts for transforming external mod

    Implements generative workflows for producing text, images, audio, and video from mixed-modal inputs.

    Pythonmlx
    Ver en GitHub↗8,254
  1. Home
  2. Artificial Intelligence & ML
  3. Audio Generation Models
  4. Multimodal Generation