vix.ing · top · new · best · stats · spec

Tu, Ming

  1. Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
    2024/07/05 by Ye Bai, Bai, Ye, Jingping Chen +106 · 27 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
  2. Efficient Neural Music Generation
    2023/05/25 by Max W. Y. Lam, Lam, Max W. Y., Qiao Tian +23 · 10 citations
    Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech and Audio Processing
  3. Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance
    2022/10/27 by Yuanzhe Chen, Ming Tu, Chen, Yuanzhe +17 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Multiple instance learning with graph neural networks
    2019/06/12 by Ming Tu, Jing Huang, Tu, Ming +5 · 2 citations
    Computer Science · #Digital Imaging for Blood Diseases #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Image Retrieval and Classification Techniques #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  5. VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
    2024/04/10 by Philip Anastassiou, Zhenyu Tang, Anastassiou, Philip +15 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  6. Reducing the Model Order of Deep Neural Networks Using Information Theory
    2016/05/16 by Tu, Ming, Berisha, Visar, Cao, Yu +1 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
  7. Simulating dysarthric speech for training data augmentation in clinical speech applications
    2018/04/27 by Jiao, Yishan, Tu, Ming, Berisha, Visar +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  8. Investigating the role of L1 in automatic pronunciation evaluation of L2 speech
    2018/07/04 by Ming Tu, Tu, Ming, Anna Grabek +5 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  9. Multi-hop Reading Comprehension across Multiple Documents by Reasoning over Heterogeneous Graphs
    2019/05/17 by Tu, Ming, Wang, Guangtao, Huang, Jing +3 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  10. Select, Answer and Explain: Interpretable Multi-hop Reading Comprehension over Multiple Documents
    2019/11/01 by Tu, Ming, Huang, Kevin, Wang, Guangtao +3 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  11. Speaker-invariant Affective Representation Learning via Adversarial Training
    2019/11/04 by Li, Haoqi, Tu, Ming, Huang, Jing +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  12. Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition
    2022/12/30 by Yukun Feng, Feng, Yukun, Ming Tu +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  13. ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction
    2025/11/11 by Yang, Shu-wen, Tu, Ming, Liu, Andy T. +5 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Signal Processing (eess.SP) #electronic engineering #information engineering