vix.ing · top · new · best · stats · spec

Keqi Deng

  1. F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching
    2024/10/09 by Yushen Chen, Zhikang Niu, Chen, Yushen +14 · 2 voices · 153 citations
    Computer Science · #Music and Audio Processing #cs.SD #eess.AS
  2. Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models
    2022/01/25 by Keqi Deng, Zehui Yang, Deng, Keqi +9 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Decoupled Structure for Improved Adaptability of End-to-End Models
    2023/08/25 by Keqi Deng, Deng, Keqi, Philip C. Woodland +1 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  4. Label-Synchronous Neural Transducer for E2E Simultaneous Speech Translation
    2024/06/06 by Keqi Deng, Philip C. Woodland, Deng, Keqi +1 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Improving CTC-based speech recognition via knowledge transferring from pre-trained language models
    2022/02/22 by Keqi Deng, Songjun Cao, Deng, Keqi +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  6. SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
    2025/04/22 by Keqi Deng, Deng, Keqi, Wenxi Chen +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. X-Translator: A Real-Time Multilingual Speaker-Aware Speech-to-Speech Translation System
    2026/07/20 by Yuxiang Zhao, Yichi Zhang, Yanjie An +10
    #eess.AS