awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Kedreamix avatar

Kedreamix/Linly-Dubbing

0
View on GitHub↗
3,048 stars·340 forks·Jupyter Notebook·apache-2.0·26 views

Linly Dubbing

Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis.

The system distinguishes itself through AI-driven lip synchronization and animation, which aligns facial expressions and mouth movements to the synthesized voiceover. It also utilizes audio source separation to isolate vocals from background music and noise, allowing for clean voice replacement while preserving original background audio.

The broader capability surface includes tools for web video downloading, timestamped speech transcription, and voice cloning. A graphical configuration interface is provided to manage the processing pipeline, select audio files, and adjust numeric parameters.

Features

  • AI Video Dubbing Tools - Coordinates a full pipeline of downloading, transcribing, translating, and synthesizing to produce dubbed videos.
  • Lip-Synced - Aligns facial expressions and mouth movements to synthetic audio to produce realistic dubbed video output.
  • Audio and Video File Transcription - Transforms spoken audio from video files into timestamped text using automated speech recognition.
  • Source Separation Tools - Isolates vocals from background music and noise to ensure clean speech replacement in dubbed content.
  • Automatic Speech Recognition - Includes a system for transcribing spoken audio into timestamped text to facilitate translation and subtitle generation.
  • Multilingual Voice Cloning Synthesizers - Provides a voice synthesis system that supports voice cloning across multiple languages for video localization.
  • Speech Synthesis - Synthesizes natural artificial speech from text with support for voice cloning and multiple languages.
  • Speech to Text Transcription - Converts spoken audio into text with precise time markers for subtitle and voiceover alignment.
  • Text-to-Speech - Generates artificial human voices from translated text to replace original audio tracks.
  • Video Localization Platforms - Localizes video content for international audiences via transcription, translation, and synthetic dubbing.
  • Language Translation Services - Provides a modular engine to translate transcribed scripts across multiple languages using swappable services.
  • Video Assembly - Combines synthetic voiceovers, original background audio, and subtitles into a final synchronized video file.
  • Script Translations - Converts transcribed scripts between languages to prepare content for synthetic audio generation.
  • Pipeline Control Panels - Ships a graphical interface for managing the dubbing pipeline, selecting audio files, and adjusting processing parameters.

Star history

Star history chart for kedreamix/linly-dubbingStar history chart for kedreamix/linly-dubbing

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does kedreamix/linly-dubbing do?

Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis.

What are the main features of kedreamix/linly-dubbing?

The main features of kedreamix/linly-dubbing are: AI Video Dubbing Tools, Lip-Synced, Audio and Video File Transcription, Source Separation Tools, Automatic Speech Recognition, Multilingual Voice Cloning Synthesizers, Speech Synthesis, Speech to Text Transcription.

Which projects share features with kedreamix/linly-dubbing?

Projects with overlapping indexed features include: elevenlabs/elevenlabs-python — This Python SDK provides a comprehensive toolkit for synthetic audio generation, voice cloning, and the development of… huanshere/videolingo — VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It… krillinai/krillinai — KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing,… abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… tmelyralab/musetalk — MuseTalk is a deep learning lip synchronization system designed to align video facial movements with audio tracks for… pluja/whishper — Whishper is a graphical user interface for transcribing audio and video files into text using the Whisper model. It…

Projects sharing features with Linly Dubbing

These projects share indexed features with Linly Dubbing. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • elevenlabs/elevenlabs-pythonelevenlabs avatar

    elevenlabs/elevenlabs-python

    2,873View on GitHub↗

    This Python SDK provides a comprehensive toolkit for synthetic audio generation, voice cloning, and the development of conversational AI agents. It enables the creation of lifelike spoken audio from text, the replication of human voices through custom cloning, and the deployment of real-time voice agents capable of interacting with external large language models. The library distinguishes itself through deep integration of conversational AI capabilities, including the design of agent personas and the execution of real-time actions via APIs. It supports professional-grade audio production thro

    Pythonartificial-intelligenceconversational-aitext-to-speech
    View on GitHub↗2,873
  • huanshere/videolingoHuanshere avatar

    Huanshere/VideoLingo

    17,498View on GitHub↗

    VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages. The system differentiates itself through a multi-step translation refinement process and a specialized natural language processing utility that segments text into single-line captions meeting broadcast standards. It also integrates synthetic voiceover generation to replace or augment original audio tracks. The projec

    Pythonai-translationdubbinglocalization
    View on GitHub↗17,498
  • krillinai/krillinaikrillinai avatar

    krillinai/KrillinAI

    9,396View on GitHub↗

    KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing, translating, and dubbing video content into multiple languages. It provides a command-line interface to chain these stages into a single production workflow, coordinating speech-to-text transcription, translation, and audio generation. The system features a translation framework that uses large language models to maintain professional terminology and natural semantics rather than literal word replacement. It includes a dubbing tool that utilizes text-to-speech and voice cloning to gene

    Godubbinglocalizationtts
    View on GitHub↗9,396
  • abus-aikorea/voice-proabus-aikorea avatar

    abus-aikorea/voice-pro

    6,255View on GitHub↗

    Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice cloning, speech recognition, and translation capabilities into a single application. At its core, the project enables users to generate natural-sounding speech from text, clone voices from short audio samples without requiring prior training data, and perform real-time speech translation across over 100 languages. The platform distinguishes itself through its integrated multimedia workflow, allowing users to download YouTube videos, extract audio, separate voice tracks, generate word

    Pythonaudiobookfaster-whispergradio
    View on GitHub↗6,255
Compare all 30 related projects→