Hu, Yuchen
- Large Language Models are Efficient Learners of Noise-Robust Speech Recognition
2024/01/19 by Yuchen Hu, Hu, Yuchen, Chen Chen +10 · 9 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Natural Language Processing Techniques
- HyPoradise: An Open Baseline for Generative Speech Recognition with Large Language Models
2023/09/27 by Chen, Chen, Hu, Yuchen, Yang, Chao-Han Huck +3 · 7 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
2025/02/05 by Yao, Jixun, Liu, Hexin, Chen, Chen +3 · 14 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition
2024/02/08 by Chen, Chen, Li, Ruizhe, Hu, Yuchen +4 · 7 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
2024/06/02 by Chen Chen, Yuchen Hu, Chen, Chen +9 · 8 citations
Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Robotics and Automated Systems #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- MEIC: Re-thinking RTL Debug Automation using LLMs
2024/05/10 by Ke Xu, Xu, Ke, Jialin Sun +11 · 7 citations
Business, Management and Accounting · Computer Science · Decision Sciences · #Advanced Database Systems and Queries #Business Process Modeling and Analysis #FOS: Computer and information sciences #Hardware Architecture (cs.AR) #Simulation Techniques and Applications #Software Engineering (cs.SE)
- Switchback Experiments under Geometric Mixing
2022/09/01 by Hu, Yuchen, Wager, Stefan · 4 citations
#Econometrics (econ.EM) #FOS: Computer and information sciences #FOS: Economics and business #Methodology (stat.ME)
- Leveraging Modality-specific Representations for Audio-visual Speech Recognition via Reinforcement Learning
2022/12/10 by Chen, Chen, Hu, Yuchen, Zhang, Qiang +3 · 4 citations
#Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
2024/02/10 by Hu, Yuchen, Chen, Chen, Yang, Chao-Han Huck +4 · 5 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Noise-robust Speech Recognition with 10 Minutes Unparalleled In-domain Data
2022/03/29 by Chen Chen, Chen, Chen, Nana Hou +7 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Audio Large Language Models Can Be Descriptive Speech Quality Evaluators
2025/01/27 by Chen Chen, Yuchen Hu, Chen, Chen +13 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Listen Again and Choose the Right Answer: A New Paradigm for Automatic Speech Recognition with Large Language Models
2024/05/16 by Yuchen Hu, Hu, Yuchen, Chen Chen +9 · 5 citations
Computer Science · #Speech Recognition and Synthesis
- Robust Zero-Shot Text-to-Speech Synthesis with Reverse Inference Optimization
2024/07/02 by Yuchen Hu, Hu, Yuchen, Chen Chen +7 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Hearing Lips in Noise: Universal Viseme-Phoneme Mapping and Transfer for Robust Audio-Visual Speech Recognition
2023/06/18 by Hu, Yuchen, Li, Ruizhe, Chen, Chen +3 · 3 citations
#Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection
2024/01/11 by Zou, Heqing, Shen, Meng, Hu, Yuchen +3 · 4 citations
#FOS: Computer and information sciences #Multimedia (cs.MM)
- VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction
2025/04/27 by Ning Wang, Bingkun Yao, Wang, Ning +11 · 8 citations
Computer Science · #Advanced Malware Detection Techniques #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Hardware Architecture (cs.AR) #Software Engineering (cs.SE) #Software Engineering Research #Software Testing and Debugging Techniques
- Average Direct and Indirect Causal Effects under Interference
2021/04/08 by Yu‐Chen Hu, Shuangning Li, Hu, Yuchen +3 · 2 citations
Mathematics · #Advanced Causal Inference Techniques #Econometrics (econ.EM) #FOS: Computer and information sciences #FOS: Economics and business #Methodology (stat.ME) #Statistical Methods and Bayesian Inference #Statistical Methods in Clinical Trials
- UVLLM: An Automated Universal RTL Verification Framework using LLMs
2024/11/25 by Yu‐Chen Hu, Hu, Yuchen, Junhao Ye +24 · 5 citations
Computer Science · #Embedded Systems Design Techniques #FOS: Computer and information sciences #Hardware Architecture (cs.AR) #Network Time Synchronization Technologies #Real-Time Systems Scheduling
- The Second Place Solution for The 4th Large-scale Video Object Segmentation Challenge--Track 3: Referring Video Object Segmentation
2022/06/24 by Cao, Leilei, Li, Zhuang, Yan, Bo +4 · 2 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM)
- Cross-Modal Global Interaction and Local Alignment for Audio-Visual Speech Recognition
2023/05/16 by Hu, Yuchen, Li, Ruizhe, Chen, Chen +3 · 2 citations
#Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- Minimax-Regret Sample Selection in Randomized Experiments
2024/03/03 by Hu, Yuchen, Zhu, Henry, Brunskill, Emma +1 · 2 citations
#Econometrics (econ.EM) #FOS: Computer and information sciences #FOS: Economics and business #Methodology (stat.ME)
- Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback
2025/04/22 by Ning Wang, Wang, Ning, Bingkun Yao +11 · 5 citations
Computer Science · #Software Engineering Research #Machine Learning and Data Classification #Software Reliability and Analysis Research
- Interactive Feature Fusion for End-to-End Noise-Robust Speech Recognition
2021/10/11 by Hu, Yuchen, Hou, Nana, Chen, Chen +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Off-Policy Evaluation in Partially Observed Markov Decision Processes under Sequential Ignorability
2021/10/24 by Hu, Yuchen, Wager, Stefan · 1 citation
#FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Methodology (stat.ME) #Statistics Theory (math.ST)
- Insights from Rights and Wrongs: A Large Language Model for Solving Assertion Failures in RTL Design
2025/03/06 by Zhou, Jie, Ji, Youshu, Wang, Ning +7 · 3 citations
#FOS: Computer and information sciences #Hardware Architecture (cs.AR)
- Interactive Audio-text Representation for Automated Audio Captioning with Contrastive Learning
2022/03/29 by Chen, Chen, Hou, Nana, Hu, Yuchen +3 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- From Concept to Practice: an Automated LLM-aided UVM Machine for RTL Verification
2025/04/28 by Junhao Ye, Yuan Hu, Ye, Junhao +18 · 3 citations
Computer Science · #Embedded Systems Design Techniques #Formal Methods in Verification #VLSI and Analog Circuit Testing
- Eeg2vec: Self-Supervised Electroencephalographic Representation Learning
2023/05/23 by Qiushi Zhu, Zhu, Qiushi, Xiaoying Zhao +9 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #EEG and Brain-Computer Interfaces #FOS: Electrical engineering #Neural dynamics and brain function #electronic engineering #information engineering
- Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
2024/04/19 by Qin, Chengwei, Xia, Wenhan, Wang, Tan +5 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Multichannel AV-wav2vec2: A Framework for Learning Multichannel Multi-Modal Speech Representation
2024/01/07 by Zhu, Qiushi, Zhang, Jie, Gu, Yu +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Attributing Response to Context: A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generation
2025/05/22 by Ruizhe Li, Li, Ruizhe, Chen Chen +9 · 2 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Domain Adaptation and Few-Shot Learning #Gaussian Processes and Bayesian Inference
- Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation Models
2024/05/23 by Yuchen Hu, Hu, Yuchen, Chen Chen +11 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Video-to-Audio Generation with Fine-grained Temporal Semantics
2024/09/23 by Y. Hu, Yu Gu, Hu, Yuchen +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Video Analysis and Summarization #electronic engineering #information engineering
- Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
2024/09/15 by Chao-Han Huck Yang, Yang, Chao-Han Huck, Taejin Park +39 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis
2024/09/11 by Helin Wang, Yu Meng, Wang, Helin +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- AnalyticKWS: Towards Exemplar-Free Analytic Class Incremental Learning for Small-footprint Keyword Spotting
2025/05/17 by Yang Xiao, Tianyi Peng, Xiao, Yang +7 · 3 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Domain Adaptation and Few-Shot Learning #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering