Zhang, Chenxu
- MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model
2023/11/27 by Zhongcong Xu, Xu, Zhongcong, Jianfeng Zhang +14 · 1 voice · 50 citations
Computer Science · #Advanced Vision and Imaging #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #cs.CV #cs.GR
- Sora Generates Videos with Stunning Geometrical Consistency
2024/02/27 by Xuanyi Li, Li, Xuanyi, Daquan Zhou +10 · 2 voices · 4 citations
Computer Science · Engineering · #Artificial Intelligence in Games #Augmented Reality Applications #Human Motion and Animation #cs.CV
- WonderHuman: Hallucinating Unseen Parts in Dynamic 3D Human Reconstruction
2025/02/03 by Zilong Wang, Wang, Zilong, Zhiyang Dou +17 · 4 voices · 1 citation
Arts and Humanities · Engineering · #Forensic Anthropology and Bioarchaeology Studies #Anatomy and Medical Technology
- Vision Mamba: A Comprehensive Survey and Taxonomy
2024/05/07 by Liu, Xiao, Zhang, Chenxu, Zhang, Lei · 15 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- DREAM-Talk: Diffusion-based Realistic Emotional Audio-driven Method for Single Image Talking Face Generation
2023/12/21 by Zhang, Chenxu, Wang, Chao, Zhang, Jianfeng +7 · 6 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- FACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute Learning
2021/08/18 by Chenxu Zhang, Yifan Zhao, Zhang, Chenxu +11 · 4 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Face recognition and analysis #Generative Adversarial Networks and Image Synthesis #Speech and Audio Processing
- X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention
2025/07/30 by Zhao, Xiaochen, Xu, Hongyi, Song, Guoxian +6 · 13 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- DiffTED: One-shot Audio-driven TED Talk Video Generation with Diffusion-based Co-speech Gestures
2024/09/11 by Steven Hogue, Hogue, Steven, Chenxu Zhang +7 · 5 citations
Computer Science · Engineering · Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Motion and Animation #Multimedia Communication and Technology #Video Analysis and Summarization
- X-Dancer: Expressive Music to Human Dance Video Generation
2025/02/24 by Chen, Zeyuan, Xu, Hongyi, Song, Guoxian +6 · 5 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- X-Dyna: Expressive Dynamic Human Image Animation
2025/01/17 by Di Chang, Chang, Di, Hongyi Xu +27 · 4 citations
Computer Science · Engineering · #3D Shape Modeling and Analysis #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Motion and Animation
- CADDreamer: CAD Object Generation from Single-view Images
2025/02/28 by Yuan Li, Li, Yuan, Cheng Lin +15 · 3 citations
Computer Science · Engineering · #3D Shape Modeling and Analysis #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Manufacturing Process and Optimization
- X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
2025/08/12 by Song, Guoxian, Xu, Hongyi, Zhao, Xiaochen +5 · 3 citations
#Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- X-Streamer: Unified Human World Modeling with Audiovisual Interaction
2025/09/25 by Xie, You, Gu, Tianpei, Li, Zenan +7 · 4 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Robust Active Speaker Detection in Noisy Environments
2024/03/27 by Siva Sai Nagender Vasireddy, Vasireddy, Siva Sai Nagender, Chenxu Zhang +5 · 1 citation
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Blind Source Separation Techniques
- Magic-Boost: Boost 3D Generation with Multi-View Conditioned Diffusion
2024/04/09 by Fan Yang, Jianfeng Zhang, Yang, Fan +16 · 1 citation
Computer Science · Engineering · #3D Shape Modeling and Analysis #Advanced Vision and Imaging #Artificial Intelligence (cs.AI) #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Empathy-R1: A Chain-of-Empathy and Reinforcement Learning Framework for Long-Form Mental Health Support
2025/09/18 by Yao, Xianrong, She, Dong, Zhang, Chenxu +5 · 1 citation
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences