vix.ing · top · new · best · stats · spec

Ming Tu

  1. Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
    2024/07/05 by Ye Bai, Bai, Ye, Jingping Chen +106 · 28 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
  2. Efficient Neural Music Generation
    2023/05/25 by Max W. Y. Lam, Qiao Tian, Lam, Max W. Y. +23 · 12 citations
    Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech and Audio Processing
  3. Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance
    2022/10/27 by Yuanzhe Chen, Ming Tu, Chen, Yuanzhe +17 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Multiple instance learning with graph neural networks
    2019/06/12 by Ming Tu, Tu, Ming, Jing Huang +5 · 3 citations
    Computer Science · #Digital Imaging for Blood Diseases #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Image Retrieval and Classification Techniques #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  5. VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing
    2024/04/10 by Philip Anastassiou, Anastassiou, Philip, Zhenyu Tang +15 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  6. Investigating the role of L1 in automatic pronunciation evaluation of L2 speech
    2018/07/04 by Ming Tu, Anna Grabek, Tu, Ming +5 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  7. Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition
    2022/12/30 by Yukun Feng, Feng, Yukun, Ming Tu +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering