awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Kedreamix avatar

Kedreamix/Linly-Dubbing

0
View on GitHub↗
3,048 星标·340 分支·Jupyter Notebook·apache-2.0·12 次浏览

Linly Dubbing

Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis.

The system distinguishes itself through AI-driven lip synchronization and animation, which aligns facial expressions and mouth movements to the synthesized voiceover. It also utilizes audio source separation to isolate vocals from background music and noise, allowing for clean voice replacement while preserving original background audio.

The broader capability surface includes tools for web video downloading, timestamped speech transcription, and voice cloning. A graphical configuration interface is provided to manage the processing pipeline, select audio files, and adjust numeric parameters.

Features

  • AI Video Dubbing Tools - Coordinates a full pipeline of downloading, transcribing, translating, and synthesizing to produce dubbed videos.
  • Lip-Synced - Aligns facial expressions and mouth movements to synthetic audio to produce realistic dubbed video output.
  • Audio and Video File Transcription - Transforms spoken audio from video files into timestamped text using automated speech recognition.
  • Source Separation Tools - Isolates vocals from background music and noise to ensure clean speech replacement in dubbed content.
  • Automatic Speech Recognition - Includes a system for transcribing spoken audio into timestamped text to facilitate translation and subtitle generation.
  • Multilingual Voice Cloning Synthesizers - Provides a voice synthesis system that supports voice cloning across multiple languages for video localization.
  • Speech Synthesis - Synthesizes natural artificial speech from text with support for voice cloning and multiple languages.
  • Speech to Text Transcription - Converts spoken audio into text with precise time markers for subtitle and voiceover alignment.
  • Text-to-Speech - Generates artificial human voices from translated text to replace original audio tracks.
  • Video Localization Platforms - Localizes video content for international audiences via transcription, translation, and synthetic dubbing.
  • Language Translation Services - Provides a modular engine to translate transcribed scripts across multiple languages using swappable services.
  • Video Assembly - Combines synthetic voiceovers, original background audio, and subtitles into a final synchronized video file.
  • Script Translations - Converts transcribed scripts between languages to prepare content for synthetic audio generation.
  • Pipeline Control Panels - Ships a graphical interface for managing the dubbing pipeline, selecting audio files, and adjusting processing parameters.

Star 历史

kedreamix/linly-dubbing 的 Star 历史图表kedreamix/linly-dubbing 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

常见问题解答

kedreamix/linly-dubbing 是做什么的?

Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis.

kedreamix/linly-dubbing 的主要功能有哪些?

kedreamix/linly-dubbing 的主要功能包括:AI Video Dubbing Tools, Lip-Synced, Audio and Video File Transcription, Source Separation Tools, Automatic Speech Recognition, Multilingual Voice Cloning Synthesizers, Speech Synthesis, Speech to Text Transcription。

kedreamix/linly-dubbing 有哪些开源替代品?

kedreamix/linly-dubbing 的开源替代品包括: elevenlabs/elevenlabs-python — This Python SDK provides a comprehensive toolkit for synthetic audio generation, voice cloning, and the development of… huanshere/videolingo — VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It… krillinai/krillinai — KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing,… abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… tmelyralab/musetalk — MuseTalk is a deep learning lip synchronization system designed to align video facial movements with audio tracks for… pluja/whishper — Whishper is a graphical user interface for transcribing audio and video files into text using the Whisper model. It…

Linly Dubbing 的开源替代方案

相似的开源项目,按与 Linly Dubbing 的功能重合度排序。
  • elevenlabs/elevenlabs-pythonelevenlabs 的头像

    elevenlabs/elevenlabs-python

    2,873在 GitHub 上查看↗

    This Python SDK provides a comprehensive toolkit for synthetic audio generation, voice cloning, and the development of conversational AI agents. It enables the creation of lifelike spoken audio from text, the replication of human voices through custom cloning, and the deployment of real-time voice agents capable of interacting with external large language models. The library distinguishes itself through deep integration of conversational AI capabilities, including the design of agent personas and the execution of real-time actions via APIs. It supports professional-grade audio production thro

    Pythonartificial-intelligenceconversational-aitext-to-speech
    在 GitHub 上查看↗2,873
  • huanshere/videolingoHuanshere 的头像

    Huanshere/VideoLingo

    17,498在 GitHub 上查看↗

    VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages. The system differentiates itself through a multi-step translation refinement process and a specialized natural language processing utility that segments text into single-line captions meeting broadcast standards. It also integrates synthetic voiceover generation to replace or augment original audio tracks. The projec

    Pythonai-translationdubbinglocalization
    在 GitHub 上查看↗17,498
  • krillinai/krillinaikrillinai 的头像

    krillinai/KrillinAI

    9,396在 GitHub 上查看↗

    KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing, translating, and dubbing video content into multiple languages. It provides a command-line interface to chain these stages into a single production workflow, coordinating speech-to-text transcription, translation, and audio generation. The system features a translation framework that uses large language models to maintain professional terminology and natural semantics rather than literal word replacement. It includes a dubbing tool that utilizes text-to-speech and voice cloning to gene

    Godubbinglocalizationtts
    在 GitHub 上查看↗9,396
  • abus-aikorea/voice-proabus-aikorea 的头像

    abus-aikorea/voice-pro

    6,255在 GitHub 上查看↗

    Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice cloning, speech recognition, and translation capabilities into a single application. At its core, the project enables users to generate natural-sounding speech from text, clone voices from short audio samples without requiring prior training data, and perform real-time speech translation across over 100 languages. The platform distinguishes itself through its integrated multimedia workflow, allowing users to download YouTube videos, extract audio, separate voice tracks, generate word

    Pythonaudiobookfaster-whispergradio
    在 GitHub 上查看↗6,255
查看 Linly Dubbing 的所有 30 个替代方案→