vix.ing · top · new · best · stats · spec

Liu, Shaoteng

  1. Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
    2024/03/27 by Yanwei Li, Li, Yanwei, Yuechen Zhang +13 · 2 voices · 43 citations
    Computer Science · #Multimodal Machine Learning Applications
  2. Video-P2P: Video Editing with Cross-attention Control
    2023/03/08 by Shaoteng Liu, Yuechen Zhang, Liu, Shaoteng +7 · 54 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications #Video Analysis and Summarization
  3. Direct Inversion: Boosting Diffusion-based Editing with 3 Lines of Code
    2023/10/02 by Xuan Ju, Ailing Zeng, Ju, Xuan +7 · 23 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cell Image Analysis Techniques #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
  4. Tent: Fully Test-time Adaptation by Entropy Minimization
    2020/06/18 by Dequan Wang, Wang, Dequan, Evan Shelhamer +7 · 8 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Multimodal Machine Learning Applications
  5. Generative Video Propagation
    2024/12/27 by Shaoteng Liu, Tianyu Wang, Liu, Shaoteng +19 · 13 citations
    Engineering · #Advanced MIMO Systems Optimization #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Telecommunications and Broadcasting Technologies
  6. RL-GPT: Integrating Reinforcement Learning and Code-as-policy
    2024/02/29 by Liu, Shaoteng, Yuan, Haoqi, Hu, Minda +5 · 4 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  7. Training-Free Efficient Video Generation via Dynamic Token Carving
    2025/05/22 by Zhang, Yuechen, Xing, Jinbo, Xia, Bin +6 · 7 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  8. EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
    2025/09/24 by Ju, Xuan, Wang, Tianyu, Zhou, Yuqian +11 · 11 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  9. Blind Robust VideoWatermarking Based on Adaptive Region Selection and Channel Reference
    2022/09/27 by Chang, Qinwei, Huang, Leichao, Liu, Shaoteng +3 · 1 citation
    #FOS: Computer and information sciences #Multimedia (cs.MM)
  10. Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and Editing
    2025/12/19 by Shilong Zhang, Zhang, Shilong, He Zhang +25 · 1 citation
    Computer Science · Engineering · #Generative Adversarial Networks and Image Synthesis #Face recognition and analysis #3D Shape Modeling and Analysis