awesome-repositories.comCategoriiBlog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Huanshere avatar

Huanshere/VideoLingo

0
View on GitHub↗
17,498 stele·1,931 fork-uri·Python·Apache-2.0·12 vizualizăridocs.videolingo.io↗

VideoLingo

VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages.

The system differentiates itself through a multi-step translation refinement process and a specialized natural language processing utility that segments text into single-line captions meeting broadcast standards. It also integrates synthetic voiceover generation to replace or augment original audio tracks.

The project covers a broad range of media processing capabilities, including automated video acquisition from external platforms, word-level timestamp alignment for subtitles, and a task sequencing system to monitor and control the localization pipeline.

Features

  • Video Localization Platforms - An integrated platform for transcribing, translating, and dubbing video media for localization.
  • AI Video Dubbing Tools - Generates synthetic voiceovers based on translated text to replace or augment original audio tracks.
  • Audio Transcription - Converts spoken video audio into precise text transcripts with word-level timing.
  • Word-Level Timestamps - Synchronizes transcribed text with precise audio timestamps for accurate subtitle timing.
  • Iterative Translation Refinement - Implements a multi-step LLM loop to translate and polish subtitles for higher accuracy.
  • Video Translation Pipelines - Implements a multi-step LLM-driven pipeline for transcribing, translating, and aligning video subtitles.
  • Speech Synthesis - Uses AI to convert translated text into artificial speech for video dubbing.
  • Automated Video Subtitling - Automatically transcribes and translates video audio to create professional, aligned captions.
  • Terminology-Aware Translation - Translates subtitles between languages using custom terminology to ensure linguistic coherence.
  • Automated Subtitle Generators - Provides an automated workflow that combines speech recognition and transcription to generate precise video subtitles in multiple languages.
  • Speech Synthesis & TTS - Converts translated subtitles into synthetic speech for video voiceovers.
  • Audio Synthesis - Generates artificial voiceovers based on translated subtitles to replace original audio.
  • Linguistic Text Segmentation - Uses NLP to segment transcribed text into readable subtitle lines following broadcast standards.
  • Text Segmentation - Divides continuous transcribed text into discrete subtitle segments based on broadcast standards.
  • Task Execution Sequencing - Orchestrates a sequence of discrete processing stages from download to final rendering.
  • Captioning Systems - Segments text into precise single-line captions that meet professional broadcast timing and layout requirements.
  • Audio and Subtitle Tools - Video translation, localization, and dubbing tool.

Istoric stele

Graficul istoricului de stele pentru huanshere/videolingoGraficul istoricului de stele pentru huanshere/videolingo

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru VideoLingo

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu VideoLingo.
  • krillinai/krillinaiAvatar krillinai

    krillinai/KrillinAI

    9,396Vezi pe GitHub↗

    KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing, translating, and dubbing video content into multiple languages. It provides a command-line interface to chain these stages into a single production workflow, coordinating speech-to-text transcription, translation, and audio generation. The system features a translation framework that uses large language models to maintain professional terminology and natural semantics rather than literal word replacement. It includes a dubbing tool that utilizes text-to-speech and voice cloning to gene

    Godubbinglocalizationtts
    Vezi pe GitHub↗9,396
  • liuzhao1225/youdub-webuiAvatar liuzhao1225

    liuzhao1225/YouDub-webui

    3,957Vezi pe GitHub↗

    YouDub-webui is a multilingual video translator and AI dubbing pipeline manager featuring a web interface for automating video translation, audio dubbing, and subtitle burning. It utilizes a GPU-accelerated media processor to speed up audio transcription and video rendering tasks. The system implements a stage-based pipeline that converts original speech into new languages while preserving background audio through audio track mixing. It supports multiple localization workflows, including automated translation and subtitle-driven dubbing using SRT files to bypass automatic transcription phases

    Python
    Vezi pe GitHub↗3,957
  • kedreamix/linly-dubbingAvatar Kedreamix

    Kedreamix/Linly-Dubbing

    3,048Vezi pe GitHub↗

    Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken content in videos into another language by coordinating speech-to-text transcription, text translation, and text-to-speech synthesis. The system distinguishes itself through AI-driven lip synchronization and animation, which aligns facial expressions and mouth movements to the synthesized voiceover. It also utilizes audio source separation to isolate vocals from background music and noise, allowing for clean voice replacement while preserving original background audio. The br

    Jupyter Notebook
    Vezi pe GitHub↗3,048
  • abus-aikorea/voice-proAvatar abus-aikorea

    abus-aikorea/voice-pro

    6,255Vezi pe GitHub↗

    Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice cloning, speech recognition, and translation capabilities into a single application. At its core, the project enables users to generate natural-sounding speech from text, clone voices from short audio samples without requiring prior training data, and perform real-time speech translation across over 100 languages. The platform distinguishes itself through its integrated multimedia workflow, allowing users to download YouTube videos, extract audio, separate voice tracks, generate word

    Pythonaudiobookfaster-whispergradio
    Vezi pe GitHub↗6,255
Vezi toate cele 30 alternative pentru VideoLingo→

Întrebări frecvente

Ce face huanshere/videolingo?

VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages.

Care sunt principalele funcționalități ale huanshere/videolingo?

Principalele funcționalități ale huanshere/videolingo sunt: Video Localization Platforms, AI Video Dubbing Tools, Audio Transcription, Word-Level Timestamps, Iterative Translation Refinement, Video Translation Pipelines, Speech Synthesis, Automated Video Subtitling.

Care sunt câteva alternative open-source pentru huanshere/videolingo?

Alternativele open-source pentru huanshere/videolingo includ: krillinai/krillinai — KrillinAI is an AI video localization pipeline and toolset designed to automate the process of transcribing,… liuzhao1225/youdub-webui — YouDub-webui is a multilingual video translator and AI dubbing pipeline manager featuring a web interface for… kedreamix/linly-dubbing — Linly-Dubbing is an automated video dubbing pipeline designed for multilingual video localization. It converts spoken… abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… tmoroney/auto-subs — Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into… chenyme/chenyme-aavt — Chenyme-AAVT is an AI-powered video transcription tool and translation platform. It converts speech from media files…