How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
CVPR 2025 Oral Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
CVPR 2025 Consistent and Controllable Image Animation with Motion Diffusion Models
CVPR 2025 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation".
CVPR 2025 Highlight🔥 Identity-Preserving Text-to-Video Generation by Frequency Decomposition
[CVPR 2025 Highlight] X-Dyna: Expressive Dynamic Human Image Animation
The main features of bytedance/x-dyna are: Computer Vision Research, Generative AI.
Projects with overlapping indexed features include: maxin-cn/cinemo — [CVPR 2025] Consistent and Controllable Image Animation with Motion Diffusion Models. taco-group/facelock — Edit Away and My Face Will not Stay: Personal Biometric Defense against Malicious Generative Editing. byteflow-ai/tokenflow — [CVPR 2025] 🔥 Official impl. of "TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation". hustvl/lightningdit — [CVPR 2025 Oral] Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models. pku-yuangroup/consisid — [CVPR 2025 Highlight🔥] Identity-Preserving Text-to-Video Generation by Frequency Decomposition. taco-group/sleepermark — [CVPR2025] We present SleeperMark, a novel framework designed to embed resilient watermarks into T2I diffusion models.