awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
yzhou359 avatar

yzhou359/MakeItTalkFork

0
View on GitHub↗
1,029 Stars·226 Forks·Jupyter Notebook·1 Aufruf

MakeItTalk

This is the code repository implementing the paper:

Features

  • Audio Driven Synthesis - Speaker-aware talking-head animation from audio.

Star-Verlauf

Star-Verlauf für yzhou359/makeittalkStar-Verlauf für yzhou359/makeittalk

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Häufig gestellte Fragen

Was macht yzhou359/makeittalk?

This is the code repository implementing the paper:

Was sind die Hauptfunktionen von yzhou359/makeittalk?

Die Hauptfunktionen von yzhou359/makeittalk sind: Audio Driven Synthesis.

Welche Open-Source-Alternativen gibt es zu yzhou359/makeittalk?

Open-Source-Alternativen zu yzhou359/makeittalk sind unter anderem: meigen-ai/infinitetalk — InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes… bytedance/latentsync — LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's… deepbrainai-research/discohead — Project Page | KoEBA Dataset. farzanehjafari1987/sedtalker — Farzaneh Jafari, Stefano Berretti, Anup Basu. fudan-generative-vision/hallo — Hallo is an audio-driven talking head generator and portrait animation framework. It synchronizes a static portrait… badtobest/echomimic — EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static…

Open-Source-Alternativen zu MakeItTalk

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit MakeItTalk.
  • meigen-ai/infinitetalkAvatar von MeiGen-AI

    MeiGen-AI/InfiniteTalk

    4,825Auf GitHub ansehen↗

    InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes realistic lip movements, head poses, and facial expressions synchronized to a spoken audio track, using either a single still image or a small set of reference video frames as the visual source. The system can produce videos of arbitrary length while maintaining temporal coherence, and it supports animating multiple subjects in a single scene. A key differentiator is the ability to coordinate multiple talking subjects through a structured JSON description, giving each independent lip

    Python
    Auf GitHub ansehen↗4,825
  • bytedance/latentsyncAvatar von bytedance

    bytedance/LatentSync

    5,806Auf GitHub ansehen↗

    LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's lip movements in a video to a target audio track. It provides a lip synchronization training framework for developing synchronization networks on custom video and audio datasets. The system utilizes a video preprocessing pipeline to clean, segment, and align face data. It includes a visual sync evaluation tool that calculates confidence scores to measure the accuracy of audio and visual alignment in generated videos. The project covers capabilities for custom synchronization

    Python
    Auf GitHub ansehen↗5,806
  • deepbrainai-research/discoheadAvatar von deepbrainai-research

    deepbrainai-research/discohead

    123Auf GitHub ansehen↗

    Project Page | KoEBA Dataset

    Python
    Auf GitHub ansehen↗123
  • badtobest/echomimicAvatar von BadToBest

    BadToBest/EchoMimic

    4,258Auf GitHub ansehen↗

    EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static reference images into dynamic talking head videos by synchronizing facial movements with audio tracks and motion drivers. The system functions as a hybrid motion synthesis engine that combines audio inputs and pose data. It utilizes a facial landmark motion controller to edit positioning markers, enabling precise synchronization and video-to-video pose transfer. The pipeline covers image-to-video animation through latent diffusion and facial landmark conditioning. This allows

    Python
    Auf GitHub ansehen↗4,258
  • Alle 30 Alternativen zu MakeItTalk anzeigen→