awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
MRzzm avatar

MRzzm/DINet

0
View on GitHub↗
1,121 stele·184 fork-uri·Python·4 vizualizări

DINet

The source code of "DINet: deformation inpainting network for realistic face visually dubbing on high resolution video."

Features

  • Audio Driven Synthesis - Deformation inpainting for realistic face dubbing on high-res video.

Istoric stele

Graficul istoricului de stele pentru mrzzm/dinetGraficul istoricului de stele pentru mrzzm/dinet

Căutare AI

Explorează mai multe repository-uri excelente

Descrie ce ai nevoie în limbaj simplu — AI-ul sortează mii de proiecte open source selectate în funcție de relevanță.

Start searching with AI

Alternative open-source pentru DINet

Proiecte open-source similare, clasificate după numărul de funcționalități comune cu DINet.
  • meigen-ai/infinitetalkAvatar MeiGen-AI

    MeiGen-AI/InfiniteTalk

    4,825Vezi pe GitHub↗

    InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes realistic lip movements, head poses, and facial expressions synchronized to a spoken audio track, using either a single still image or a small set of reference video frames as the visual source. The system can produce videos of arbitrary length while maintaining temporal coherence, and it supports animating multiple subjects in a single scene. A key differentiator is the ability to coordinate multiple talking subjects through a structured JSON description, giving each independent lip

    Python
    Vezi pe GitHub↗4,825
  • bytedance/latentsyncAvatar bytedance

    bytedance/LatentSync

    5,806Vezi pe GitHub↗

    LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's lip movements in a video to a target audio track. It provides a lip synchronization training framework for developing synchronization networks on custom video and audio datasets. The system utilizes a video preprocessing pipeline to clean, segment, and align face data. It includes a visual sync evaluation tool that calculates confidence scores to measure the accuracy of audio and visual alignment in generated videos. The project covers capabilities for custom synchronization

    Python
    Vezi pe GitHub↗5,806
  • deepbrainai-research/discoheadAvatar deepbrainai-research

    deepbrainai-research/discohead

    123Vezi pe GitHub↗

    Project Page | KoEBA Dataset

    Python
    Vezi pe GitHub↗123
  • badtobest/echomimicAvatar BadToBest

    BadToBest/EchoMimic

    4,258Vezi pe GitHub↗

    EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static reference images into dynamic talking head videos by synchronizing facial movements with audio tracks and motion drivers. The system functions as a hybrid motion synthesis engine that combines audio inputs and pose data. It utilizes a facial landmark motion controller to edit positioning markers, enabling precise synchronization and video-to-video pose transfer. The pipeline covers image-to-video animation through latent diffusion and facial landmark conditioning. This allows

    Python
    Vezi pe GitHub↗4,258
Vezi toate cele 30 alternative pentru DINet→

Întrebări frecvente

Ce face mrzzm/dinet?

The source code of "DINet: deformation inpainting network for realistic face visually dubbing on high resolution video."

Care sunt principalele funcționalități ale mrzzm/dinet?

Principalele funcționalități ale mrzzm/dinet sunt: Audio Driven Synthesis.

Care sunt câteva alternative open-source pentru mrzzm/dinet?

Alternativele open-source pentru mrzzm/dinet includ: meigen-ai/infinitetalk — InfiniteTalk is an open-source system for generating talking head videos driven by audio input. It synthesizes… bytedance/latentsync — LatentSync is an audio-driven video generator and latent diffusion lip sync model designed to synchronize a speaker's… deepbrainai-research/discohead — Project Page | KoEBA Dataset. farzanehjafari1987/sedtalker — Farzaneh Jafari, Stefano Berretti, Anup Basu. fudan-generative-vision/hallo — Hallo is an audio-driven talking head generator and portrait animation framework. It synchronizes a static portrait… badtobest/echomimic — EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static…