awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
antgroup avatar

antgroup/echomimic

0
View on GitHub↗
4,255 estrellas·463 forks·Python·Apache-2.0·6 vistasantgroup.github.io/ai/echomimic↗

Echomimic

EchoMimic is a multimodal human animation framework and diffusion-based video generator. It produces lifelike facial and semi-body animations of a reference image by synthesizing motion and appearance from various source data.

The system enables portrait animation driven by audio, pose sequences, or driver videos. It features a landmark conditioning tool that allows for the precise control of facial movements by modifying specific landmark points.

The framework covers multi-modal motion synthesis and the synchronization of reference images to match the physical movements of a target driver. This includes the ability to transform audio signals into facial pose parameters to drive generated video frames.

Features

  • Portrait Animation Engines - Generates lifelike facial animations by syncing a reference image to a provided audio track.
  • Latent Diffusion Frame Synthesizers - Uses a large-parameter neural network and diffusion priors to synthesize high-fidelity video frames.
  • Video Diffusion Models - Uses a diffusion-based generative model to produce high-quality video sequences from multimodal source data.
  • Image-Conditioned Video Generators - Injects visual features from a static reference image to maintain identity and texture consistency in generated videos.
  • Multi-Modal Conditioned Synthesis - Animates portraits using a single model driven by diverse inputs such as audio, pose sequences, or reference videos.
  • Multi-Modal Animation Frameworks - Executes human animation tasks across audio and pose inputs using a unified high-parameter model.
  • Multi-Modal Motion Drivers - Combines audio and pose data into a unified latent space to control subject appearance and motion.
  • Image-to-Video Generation - Synchronizes a reference image to match the physical movements of a target driver video.
  • Facial Animation Models - Provides a deep learning model for generating lifelike facial animations synchronized to audio input.
  • Audio-to-Motion Embeddings - Implements a neural mapping that transforms raw audio signals into facial pose parameters for video generation.
  • Pose-Based Animation Alignment - Creates video animations of a reference image driven by specific pose sequences or driver videos.
  • Animation Drivers - Framework for animating human figures using diverse inputs such as audio, pose sequences, or driver videos.
  • Human Image and Video Generation - Creates semi-body human videos that synchronize a reference image with the motion of a driver video.
  • Human Motion Synthesis - Produces semi-body human animations with consistent movement and visual quality.
  • Portrait Animation Tools - Generates portrait animations from audio inputs using editable landmark conditioning to drive facial movement.
  • Landmark Editing Tools - Provides a tool for precise facial movement control by modifying specific landmark points.
  • Latent Space Manipulations - Performs motion updates within a compressed latent representation to ensure temporal stability and efficiency.

Historial de estrellas

Gráfico del historial de estrellas de antgroup/echomimicGráfico del historial de estrellas de antgroup/echomimic

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Alternativas open-source a Echomimic

Proyectos open-source similares, clasificados según cuántas características comparten con Echomimic.
  • fudan-generative-vision/halloAvatar de fudan-generative-vision

    fudan-generative-vision/hallo

    8,644Ver en GitHub↗

    Hallo is an audio-driven talking head generator and portrait animation framework. It synchronizes a static portrait image with an audio file to produce realistic talking head videos by mapping audio spectral features to facial expressions and lip movements. The system utilizes a diffusion video synthesis model that employs iterative denoising and latent representations to generate temporally consistent video frames. It incorporates identity-preserving feature extraction and latent space motion modeling to maintain visual consistency and control facial poses. The toolkit provides capabilities

    Pythonface-animationimage-animationvideo-animation
    Ver en GitHub↗8,644
  • badtobest/echomimicAvatar de BadToBest

    BadToBest/EchoMimic

    4,258Ver en GitHub↗

    EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static reference images into dynamic talking head videos by synchronizing facial movements with audio tracks and motion drivers. The system functions as a hybrid motion synthesis engine that combines audio inputs and pose data. It utilizes a facial landmark motion controller to edit positioning markers, enabling precise synchronization and video-to-video pose transfer. The pipeline covers image-to-video animation through latent diffusion and facial landmark conditioning. This allows

    Python
    Ver en GitHub↗4,258
  • fudan-generative-vision/hallo2Avatar de fudan-generative-vision

    fudan-generative-vision/hallo2

    3,713Ver en GitHub↗

    Hallo2 is an AI video generation tool and audio-driven portrait animation framework designed to transform static images into speaking videos. It functions as a portrait image animator that synchronizes a single photo with an audio track to produce high-resolution talking head videos. The system includes a distributed animation trainer for fine-tuning deep learning models using custom datasets and distributed computing resources. It employs hierarchical video generation and temporal consistency modeling to produce long-form character animations that remain stable over extended durations. The

    Python
    Ver en GitHub↗3,713
  • hvision-nku/storydiffusionAvatar de HVision-NKU

    HVision-NKU/StoryDiffusion

    6,430Ver en GitHub↗

    StoryDiffusion is a generative AI system designed for consistent character image and video generation. It utilizes a pluggable cross-attention module to inject shared character representations into pretrained diffusion models, allowing for visual identity stability across multiple images and scenes without retraining the base model. The project features a video generation pipeline that produces temporally coherent sequences from text prompts or condition images. It employs a latent space motion interpolator to predict intermediate frames and semantic motion, enabling long-range video generati

    Jupyter Notebook
    Ver en GitHub↗6,430
Ver las 30 alternativas a Echomimic→

Preguntas frecuentes

¿Qué hace antgroup/echomimic?

EchoMimic is a multimodal human animation framework and diffusion-based video generator. It produces lifelike facial and semi-body animations of a reference image by synthesizing motion and appearance from various source data.

¿Cuáles son las características principales de antgroup/echomimic?

Las características principales de antgroup/echomimic son: Portrait Animation Engines, Latent Diffusion Frame Synthesizers, Video Diffusion Models, Image-Conditioned Video Generators, Multi-Modal Conditioned Synthesis, Multi-Modal Animation Frameworks, Multi-Modal Motion Drivers, Image-to-Video Generation.

¿Qué alternativas de código abierto existen para antgroup/echomimic?

Las alternativas de código abierto para antgroup/echomimic incluyen: fudan-generative-vision/hallo — Hallo is an audio-driven talking head generator and portrait animation framework. It synchronizes a static portrait… badtobest/echomimic — EchoMimic is an audio-driven portrait animation framework and latent diffusion video generator. It transforms static… fudan-generative-vision/hallo2 — Hallo2 is an AI video generation tool and audio-driven portrait animation framework designed to transform static… hvision-nku/storydiffusion — StoryDiffusion is a generative AI system designed for consistent character image and video generation. It utilizes a… aigc-apps/sd-webui-easyphoto — This project is a Stable Diffusion WebUI extension that provides a graphical interface for personalized portrait… humanaigc/animateanyone — AnimateAnyone is an appearance-preserving video synthesizer designed for character animation from a single static…