Xue, Hongfei
- Fusing Global and Local Features for Generalized AI-Synthesized Image Detection
2022/03/26 by Yiguang Ju, Ju, Yan, Shan Jia +9 · 17 citations
Computer Science · #Digital Media Forensic Detection #Generative Adversarial Networks and Image Synthesis #Anomaly Detection Techniques and Applications
- E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models
2023/12/31 by Hongfei Xue, Yuhao Liang, Xue, Hongfei +10 · 5 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
2024/10/14 by Ekkasit Pinyoanuntapong, Pinyoanuntapong, Ekkasit, Muhammad Usama Saleem +16 · 5 citations
Engineering · Computer Science · #Human Motion and Animation #Augmented Reality Applications #Interactive and Immersive Displays
- Unveiling the Potential of LLM-Based ASR on Chinese Open-Source Datasets
2024/05/03 by Xuelong Geng, Tianyi Xu, Geng, Xuelong +21 · 6 citations
Computer Science · #Natural Language Processing Techniques
- HOIGPT: Learning Long Sequence Hand-Object Interaction with Language Models
2025/03/24 by Huang, Mingzhen, Chu, Fu-Jen, Tekin, Bugra +10 · 5 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- mmCooper: A Multi-agent Multi-stage Communication-efficient and Collaboration-robust Cooperative Perception Framework
2025/01/21 by Liu, Bingyi, Teng, Jian, Xue, Hongfei +4 · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Argus: Multi-View Egocentric Human Mesh Reconstruction Based on Stripped-Down Wearable mmWave Add-on
2024/11/01 by Duan, Di, Lyu, Shengzhe, Yuan, Mu +5 · 3 citations
#C.3 #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC)
- BioPose: Biomechanically-accurate 3D Pose Estimation from Monocular Videos
2025/01/14 by Koleini, Farnoosh, Saleem, Muhammad Usama, Wang, Pu +3 · 4 citations
#Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Ideal-LLM: Integrating Dual Encoders and Language-Adapted LLM for Multilingual Speech-to-Text
2024/09/17 by Xue, Hongfei, Ren, Wei, Geng, Xuelong +6 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue
2025/08/13 by Geng, Xuelong, Shao, Qijie, Xue, Hongfei +20 · 5 citations
#FOS: Computer and information sciences #Sound (cs.SD)
- Hearing More with Less: Multi-Modal Retrieval-and-Selection Augmented Conversational LLM-Based ASR
2025/08/02 by Bingshen Mu, Mu, Bingshen, Hexin Liu +7 · 2 citations
Computer Science · Medicine · #FOS: Computer and information sciences #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #Voice and Speech Disorders
- WenetSpeech-Chuan: A Large-Scale Sichuanese Corpus with Rich Annotation for Dialectal Speech Processing
2025/09/22 by Dai, Yuhang, Zhang, Ziyu, Wang, Shuai +13 · 3 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences #Sound (cs.SD)
- WenetSpeech-Yue: A Large-scale Cantonese Speech Corpus with Multi-dimensional Annotation
2025/09/04 by Longhao Li, Zhao Guo, Li, Longhao +33 · 3 citations
Computer Science · #FOS: Computer and information sciences #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling