Yin-Dong Zheng
- VideoLLM: Modeling Video Sequence with Large Language Models
2023/05/22 by Chen Guo, Yin-Dong Zheng, Chen, Guo +19 · 24 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Natural Language Processing Techniques
- InternVideo-Ego4D: A Pack of Champion Solutions to Ego4D Challenges
2022/11/17 by Chen Guo, Sen Xing, Chen, Guo +39 · 5 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications
- DCAN: Improving Temporal Action Detection via Dual Context Aggregation
2021/12/07 by Guo Chen, Yin-Dong Zheng, Chen, Guo +5 · 2 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Video Surveillance and Tracking Methods
- BasicTAD: an Astounding RGB-Only Baseline for Temporal Action Detection
2022/05/05 by Min Yang, Yang, Min, Guo Chen +7 · 1 citation
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Video Surveillance and Tracking Methods