vix.ing · top · new · best · stats · spec

Wang, Hongfa

  1. HunyuanVideo: A Systematic Framework For Large Video Generative Models
    2024/12/03 by Weijie Kong, Kong, Weijie, Qi Tian +106 · 1 voice · 375 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Video Analysis and Summarization #cs.CV
  2. Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
    2024/06/04 by Ma, Yue, Liu, Hongyu, Wang, Hongfa +8 · 45 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  3. Egocentric Video-Language Pretraining
    2022/06/03 by Kevin Qinghong Lin, Lin, Kevin Qinghong, Alex Jinpeng Wang +29 · 19 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #Cancer-related molecular mechanisms research #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications
  4. MAP: Multimodal Uncertainty-Aware Vision-Language Pre-training Model
    2022/10/11 by Ji, Yatai, Wang, Junjie, Gong, Yuan +6 · 9 citations
    #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM)
  5. Controllable Video Generation: A Survey
    2025/07/22 by Yue Ma, Ma, Yue, Kun Feng +36 · 32 citations
    Computer Science · #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Graphics (cs.GR) #Video Analysis and Summarization #Video Coding and Compression Technologies
  6. Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation
    2024/09/02 by Chen, Qihua, Ma, Yue, Wang, Hongfa +7 · 12 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling
    2024/06/05 by Xue, Jingyun, Wang, Hongfa, Tian, Qi +10 · 8 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  8. CoHD: A Counting-Aware Hierarchical Decoding Framework for Generalized Referring Expression Segmentation
    2024/05/24 by Zhuoyan Luo, Yinghao Wu, Luo, Zhuoyan +11 · 7 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling
  9. Global and Local Semantic Completion Learning for Vision-Language Pre-training
    2023/06/12 by Rong-Cheng Tu, Yatai Ji, Tu, Rong-Cheng +15 · 3 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  10. Tencent Text-Video Retrieval: Hierarchical Cross-Modal Interactions with Multi-Level Representations
    2022/04/07 by Jiang, Jie, Min, Shaobo, Kong, Weijie +4 · 2 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  11. Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts
    2024/03/13 by Ma, Yue, He, Yingqing, Wang, Hongfa +8 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  12. Inverse-like Antagonistic Scene Text Spotting via Reading-Order Estimation and Dynamic Sampling
    2024/01/08 by Zhang, Shi-Xue, Yang, Chun, Zhu, Xiaobin +3 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  13. Deep Relational Reasoning Graph Network for Arbitrary Shape Text Detection
    2020/03/17 by Zhang, Shi-Xue, Zhu, Xiaobin, Hou, Jie-Bo +4 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  14. Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
    2025/09/20 by Yue Ma, Zexuan Yan, Ma, Yue +25 · 14 citations
    Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Motion and Animation #Video Analysis and Summarization
  15. Egocentric Video-Language Pretraining @ Ego4D Challenge 2022
    2022/07/04 by Kevin Qinghong Lin, Lin, Kevin Qinghong, Alex Jinpeng Wang +29 · 1 citation
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  16. VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
    2025/05/29 by Zhang, Shi-Xue, Wang, Hongfa, Huang, Duojun +3 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  17. Object-AVEdit: An Object-level Audio-Visual Editing Model
    2025/09/27 by Fu, Youquan, Si, Ruiyang, Wang, Hongfa +6 · 7 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
  18. Seeing What You Miss: Vision-Language Pre-training with Semantic Completion Learning
    2022/11/24 by Yatai Ji, Ji, Yatai, Rong-Cheng Tu +15 · 1 citation
    Computer Science · #Advanced Image and Video Retrieval Techniques #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimedia (cs.MM) #Multimodal Machine Learning Applications
  19. BalanceBenchmark: A Survey for Multimodal Imbalance Learning
    2025/02/15 by Shaoxuan Xu, Xu, Shaoxuan, Cui, Menglu +6 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Imbalanced Data Classification Techniques #Machine Learning (cs.LG)
  20. LayoutDiT: Exploring Content-Graphic Balance in Layout Generation with Diffusion Transformer
    2024/07/21 by Li Yu, Li, Yu, Yifan Chen +14 · 1 citation
    Computer Science · #Video Analysis and Summarization #Semantic Web and Ontologies #Image Retrieval and Classification Techniques
  21. Efficient Quantification of Multimodal Interaction at Sample Level
    2025/06/08 by Yang, Zequn, Wang, Hongfa, Hu, Di · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)