awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

2 个仓库

Awesome GitHub RepositoriesAudio Gap Infilling

Neural processes for predicting and restoring missing segments of audio recordings.

Distinct from Audio Processing: Specifically targets the restoration of missing audio segments, whereas general audio processing covers a broader range of manipulations.

Explore 2 awesome GitHub repositories matching graphics & multimedia · Audio Gap Infilling. Refine with filters or upvote what's useful.

Awesome Audio Gap Infilling GitHub Repositories

用 AI 发现最棒的仓库。我们将通过 AI 为您搜索最匹配的仓库。
  • aigc-audio/audiogptAIGC-Audio 的头像

    AIGC-Audio/AudioGPT

    10,174在 GitHub 上查看↗

    AudioGPT is an LLM-driven audio framework and processing suite that uses large language models to orchestrate neural audio pipelines. It functions as a multimodal audio generator and processing system, integrating a collection of pretrained models to handle speech synthesis, sound generation, and audio manipulation. The system is distinguished by its ability to generate audio from diverse inputs, including text and images, and its capacity to produce synchronized talking head videos. It also operates as a neural speech translator, converting spoken language between different tongues while pre

    Restores missing segments of sound recordings by predicting and inserting the absent audio data.

    Pythonaudiogptmusic
    在 GitHub 上查看↗10,174
  • jasonppy/voicecraftjasonppy 的头像

    jasonppy/VoiceCraft

    8,500在 GitHub 上查看↗

    VoiceCraft is a neural speech generation and manipulation system consisting of a text-to-speech system, a voice cloning tool, and an audio inpainting engine. It uses a large language model approach to synthesize high-fidelity audio from text and replicate speaker identities. The system provides zero-shot voice cloning and speech editing capabilities, allowing users to modify spoken content within existing recordings. This includes an audio inpainting engine that replaces specific sections of audio with new speech while preserving the original acoustic characteristics and speaker identity. Th

    Provides a neural engine for predicting and restoring missing audio segments to modify spoken content.

    Jupyter Notebook
    在 GitHub 上查看↗8,500
  1. Home
  2. Graphics & Multimedia
  3. Audio & Music
  4. Audio Processing
  5. Audio Gap Infilling