Tu, Ming
- Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
2024/07/05 by Ye Bai, Bai, Ye, Jingping Chen +106 · 27 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
- Efficient Neural Music Generation
2023/05/25 by Max W. Y. Lam, Lam, Max W. Y., Qiao Tian +23 · 10 citations
Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech and Audio Processing
- Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance
2022/10/27 by Yuanzhe Chen, Ming Tu, Chen, Yuanzhe +17 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Multiple instance learning with graph neural networks
2019/06/12 by Ming Tu, Jing Huang, Tu, Ming +5 · 2 citations
Computer Science · #Digital Imaging for Blood Diseases #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Image Retrieval and Classification Techniques #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
2024/04/10 by Philip Anastassiou, Zhenyu Tang, Anastassiou, Philip +15 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Reducing the Model Order of Deep Neural Networks Using Information Theory
2016/05/16 by Tu, Ming, Berisha, Visar, Cao, Yu +1 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
- Simulating dysarthric speech for training data augmentation in clinical speech applications
2018/04/27 by Jiao, Yishan, Tu, Ming, Berisha, Visar +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
- Investigating the role of L1 in automatic pronunciation evaluation of L2 speech
2018/07/04 by Ming Tu, Tu, Ming, Anna Grabek +5 · 1 citation
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Multi-hop Reading Comprehension across Multiple Documents by Reasoning over Heterogeneous Graphs
2019/05/17 by Tu, Ming, Wang, Guangtao, Huang, Jing +3 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Select, Answer and Explain: Interpretable Multi-hop Reading Comprehension over Multiple Documents
2019/11/01 by Tu, Ming, Huang, Kevin, Wang, Guangtao +3 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Speaker-invariant Affective Representation Learning via Adversarial Training
2019/11/04 by Li, Haoqi, Tu, Ming, Huang, Jing +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition
2022/12/30 by Yukun Feng, Feng, Yukun, Ming Tu +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction
2025/11/11 by Yang, Shu-wen, Tu, Ming, Liu, Andy T. +5 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Signal Processing (eess.SP) #electronic engineering #information engineering