Mo, Sicheng
- FreeControl: Training-Free Spatial Control of Any Text-to-Image Diffusion Model with Any Condition
2023/12/12 by Sicheng Mo, Fangzhou Mu, Mo, Sicheng +11 · 1 voice · 16 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis
- Ctrl-X: Controlling Structure and Appearance for Text-To-Image Generation Without Guidance
2024/06/11 by Kuan Heng Lin, Sicheng Mo, Lin, Kuan Heng +7 · 1 voice · 6 citations
Computer Science · Engineering · #Augmented Reality Applications #Human Motion and Animation
- SnAG: Scalable and Accurate Video Grounding
2024/04/02 by Fangzhou Mu, Sicheng Mo, Mu, Fangzhou +3 · 10 citations
Computer Science · #Advanced Neural Network Applications #Multimodal Machine Learning Applications #Advanced Image and Video Retrieval Techniques
- SimGen: Simulator-conditioned Driving Scene Generation
2024/06/13 by Zhou, Yunsong, Simon, Michael, Peng, Zhenghao +4 · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- X-Fusion: Introducing New Modality to Frozen Large Language Models
2025/04/29 by Mo, Sicheng, Nguyen, Thao, Huang, Xun +9 · 6 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Physics to the Rescue: Deep Non-line-of-sight Reconstruction for High-speed Imaging
2022/05/03 by Mu, Fangzhou, Mo, Sicheng, Peng, Jiayong +5 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #electronic engineering #information engineering
- A Simple Transformer-Based Model for Ego4D Natural Language Queries Challenge
2022/11/16 by Sicheng Mo, Fangzhou Mu, Mo, Sicheng +3 · 1 citation
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Video Analysis and Summarization