awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
liuzhao1225 avatar

liuzhao1225/YouDub-webui

0
View on GitHub↗
3,957 stars·408 forks·Python·11 vues

YouDub Webui

YouDub-webui is a multilingual video translator and AI dubbing pipeline manager featuring a web interface for automating video translation, audio dubbing, and subtitle burning. It utilizes a GPU-accelerated media processor to speed up audio transcription and video rendering tasks.

The system implements a stage-based pipeline that converts original speech into new languages while preserving background audio through audio track mixing. It supports multiple localization workflows, including automated translation and subtitle-driven dubbing using SRT files to bypass automatic transcription phases.

The project includes a task management system for tracking real-time pipeline progress and execution logs. This framework allows users to resume failed localization jobs from the last unsuccessful stage by reusing cached outputs from completed steps.

Users can manage API credentials, network concurrency settings, and external service integrations directly through the web interface.

Features

  • Audio-Visual Translation - Provides a comprehensive system for transcribing audio and generating new localized voiceovers while preserving background sounds.
  • Video Localization Platforms - Provides an integrated platform for transcribing, translating, and dubbing videos automatically.
  • AI Video Dubbing Tools - Implements a workflow to replace original video audio with synthetic translated voiceovers while preserving background sounds.
  • Subtitle-Driven Dubbing - Generates dubbed voiceovers and burns subtitles into videos using SRT files instead of automatic transcription.
  • Subtitle-Driven Audio Synthesis - Generates dubbed audio by utilizing the timing and content of provided SRT files.
  • GPU-Accelerated Inference - Uses GPU-accelerated inference to optimize the speed of audio transcription and voice synthesis models.
  • Multi-Platform Video Translation - Processes video localization from multiple sources, including platform links and local file uploads.
  • Subtitle-Based Localization - Allows users to localize uploaded videos using provided subtitles to bypass transcription and translation phases.
  • Localization Pipeline Stages - Structures the localization process into a sequence of discrete, cacheable steps from transcription to final rendering.
  • Audio Mixing - Combines generated synthetic speech with the original background audio to maintain a natural sonic atmosphere.
  • Audio Synthesis - Transcribes original speech and generates new translated voiceovers for video localization.
  • Hardware-Accelerated Media Processors - Implements a processing engine that leverages GPU acceleration to speed up audio transcription and video rendering tasks.
  • Video Translation Pipelines - Manages the sequential pipeline of transcribing, translating, and aligning audio for video localization.
  • Intermediate Output Caching - Caches intermediate transcription and translation results to allow tasks to resume without repeating expensive computations.
  • Job State Persistence - Persists real-time progress and execution logs in a database to monitor long-running asynchronous localization jobs.
  • Task & Job Management - Manages the execution, monitoring, and re-running of video localization tasks.
  • Task Pause and Resume Controls - Allows localization jobs to resume from the last successful stage using cached intermediate outputs.
  • Task Progress Monitors - Provides utilities for tracking the real-time progress, status, and timing of video localization tasks.
  • Real-Time Monitoring Dashboards - Ships a web interface for monitoring the real-time status, logs, and stage-specific durations of the processing pipeline.

Historique des stars

Graphique de l'historique des stars pour liuzhao1225/youdub-webuiGraphique de l'historique des stars pour liuzhao1225/youdub-webui

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Questions fréquentes

Que fait liuzhao1225/youdub-webui ?

YouDub-webui is a multilingual video translator and AI dubbing pipeline manager featuring a web interface for automating video translation, audio dubbing, and subtitle burning. It utilizes a GPU-accelerated media processor to speed up audio transcription and video rendering tasks.

Quelles sont les fonctionnalités principales de liuzhao1225/youdub-webui ?

Les fonctionnalités principales de liuzhao1225/youdub-webui sont : Audio-Visual Translation, Video Localization Platforms, AI Video Dubbing Tools, Subtitle-Driven Dubbing, Subtitle-Driven Audio Synthesis, GPU-Accelerated Inference, Multi-Platform Video Translation, Subtitle-Based Localization.

Quelles sont les alternatives open-source à liuzhao1225/youdub-webui ?

Les alternatives open-source à liuzhao1225/youdub-webui incluent : huanshere/videolingo — VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It… krillinai/krillinai — KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing,… kedreamix/linly-dubbing — Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken… lmms/lmms — LMMS is a digital audio workstation and MIDI sequencer designed for composing, arranging, and mixing music. It… tmelyralab/musetalk — MuseTalk is a deep learning lip synchronization system designed to align video facial movements with audio tracks for… naudio/naudio — NAudio is a .NET audio library that provides playback, recording, format conversion, and signal processing…

Alternatives open source à YouDub Webui

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec YouDub Webui.
  • huanshere/videolingoAvatar de Huanshere

    Huanshere/VideoLingo

    17,498Voir sur GitHub↗

    VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages. The system differentiates itself through a multi-step translation refinement process and a specialized natural language processing utility that segments text into single-line captions meeting broadcast standards. It also integrates synthetic voiceover generation to replace or augment original audio tracks. The projec

    Pythonai-translationdubbinglocalization
    Voir sur GitHub↗17,498
  • krillinai/krillinaiAvatar de krillinai

    krillinai/KrillinAI

    9,396Voir sur GitHub↗

    KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing, translating, and dubbing video content into multiple languages. It provides a command-line interface to chain these stages into a single production workflow, coordinating speech-to-text transcription, translation, and audio generation. The system features a translation framework that uses large language models to maintain professional terminology and natural semantics rather than literal word replacement. It includes a dubbing tool that utilizes text-to-speech and voice cloning to gene

    Godubbinglocalizationtts
    Voir sur GitHub↗9,396
  • kedreamix/linly-dubbingAvatar de Kedreamix

    Kedreamix/Linly-Dubbing

    3,048Voir sur GitHub↗

    Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis. The system distinguishes itself through AI-driven lip synchronization and animation, which aligns facial expressions and mouth movements to the synthesized voiceover. It also utilizes audio source separation to isolate vocals from background music and noise, allowing for clean voice replacement while preserving original background audio. The br

    Jupyter Notebook
    Voir sur GitHub↗3,048
  • lmms/lmmsAvatar de LMMS

    LMMS/lmms

    10,005Voir sur GitHub↗

    LMMS is a digital audio workstation and MIDI sequencer designed for composing, arranging, and mixing music. It functions as a comprehensive production environment that integrates a MIDI sequencer, a sample-based synthesizer, and an audio mixing console. The project distinguishes itself through a versatile synthesis engine that includes additive synthesis, wavetable generation, and emulations of vintage hardware such as NES audio and FM chips. It also serves as a VST plugin host, allowing for the integration of third-party virtual instruments and audio effects via a standardized interface. Be

    C++dawhacktoberfestmidi
    Voir sur GitHub↗10,005
  • Voir les 30 alternatives à YouDub Webui→