Bingkun Huang
- VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking
2023/03/29 by Limin Wang, Bingkun Huang, Wang, Limin +13 · 95 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- InternVideo: General Video Foundation Models via Generative and Discriminative Learning
2022/12/06 by Yi Wang, Kunchang Li, Wang, Yi +31 · 84 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications
- MGMAE: Motion Guided Masking for Video Masked Autoencoding
2023/08/21 by Bingkun Huang, Zhiyu Zhao, Huang, Bingkun +7 · 8 citations
Computer Science · #Advanced Image Processing Techniques #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG)
- InternVideo-Ego4D: A Pack of Champion Solutions to Ego4D Challenges
2022/11/17 by Chen Guo, Chen, Guo, Sen Xing +39 · 5 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications