awesome-repositories.comKategorienBlog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
deepbrainai-research avatar

deepbrainai-research/discohead

0
View on GitHub↗
123 Stars·12 Forks·Python·5 Aufrufe

Discohead

Project Page | KoEBA Dataset

Features

  • Audio Driven Synthesis - Disentangled control of head pose and facial expressions.

Star-Verlauf

Star-Verlauf für deepbrainai-research/discoheadStar-Verlauf für deepbrainai-research/discohead

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu Discohead

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Discohead.
  • meigen-ai/infinitetalkAvatar von MeiGen-AI

    MeiGen-AI/InfiniteTalk

    4,825Auf GitHub ansehen↗

    InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes realistic lip movements, head poses, and facial expressions synchronized to a spoken audio track, using either a single still image or a small set of reference video frames as the visual source. The system can produce videos of arbitrary length while maintaining temporal coherence, and it supports animating multiple subjects in a single scene. A key differentiator is the ability to coordinate multiple talking subjects through a structured JSON description, giving each independent lip

    Python
    Auf GitHub ansehen↗4,825
  • bytedance/latentsyncAvatar von bytedance

    bytedance/LatentSync

    5,806Auf GitHub ansehen↗

    LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's lip movements in a video to a target audio track. It provides a lip synchronization training framework for developing synchronization networks on custom video and audio datasets. The system utilizes a video preprocessing pipeline to clean, segment, and align face data. It includes a visual sync evaluation tool that calculates confidence scores to measure the accuracy of audio and visual alignment in generated videos. The project covers capabilities for custom synchronization

    Python
    Auf GitHub ansehen↗5,806
  • farzanehjafari1987/sedtalkerAvatar von FarzanehJafari1987

    FarzanehJafari1987/SEDTalker

    3Auf GitHub ansehen↗

    Farzaneh Jafari, Stefano Berretti, Anup Basu

    Python
    Auf GitHub ansehen↗3
  • badtobest/echomimicAvatar von BadToBest

    BadToBest/EchoMimic

    4,258Auf GitHub ansehen↗

    EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static reference images into dynamic talking head videos by synchronizing facial movements with audio tracks and motion drivers. The system functions as a hybrid motion synthesis engine that combines audio inputs and pose data. It utilizes a facial landmark motion controller to edit positioning markers, enabling precise synchronization and video-to-video pose transfer. The pipeline covers image-to-video animation through latent diffusion and facial landmark conditioning. This allows

    Python
    Auf GitHub ansehen↗4,258
Alle 30 Alternativen zu Discohead anzeigen→

Häufig gestellte Fragen

Was macht deepbrainai-research/discohead?

Project Page | KoEBA Dataset

Was sind die Hauptfunktionen von deepbrainai-research/discohead?

Die Hauptfunktionen von deepbrainai-research/discohead sind: Audio Driven Synthesis.

Welche Open-Source-Alternativen gibt es zu deepbrainai-research/discohead?

Open-Source-Alternativen zu deepbrainai-research/discohead sind unter anderem: meigen-ai/infinitetalk — InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes… bytedance/latentsync — LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's… farzanehjafari1987/sedtalker — Farzaneh Jafari, Stefano Berretti, Anup Basu. fudan-generative-vision/hallo — Hallo is an audio-driven talking head generator and portrait animation framework. It synchronizes a static portrait… fudan-generative-vision/hallo2 — Hallo2 is an AI video generation tool and audio-driven portrait animation framework designed to transform static… badtobest/echomimic — EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static…