vix.ing · top · new · best · stats · spec

Doddipatla, Rama

  1. Teacher-Student MixIT for Unsupervised and Semi-supervised Speech Separation
    2021/06/15 by Jisi Zhang, Cătălin Zorilă, Zhang, Jisi +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. A study on cross-corpus speech emotion recognition and data augmentation
    2022/01/10 by Norbert Braunschweiler, Braunschweiler, Norbert, Rama Doddipatla +5 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. An Investigation into the Effectiveness of Enhancement in ASR Training and Test for CHiME-5 Dinner Party Transcription
    2019/09/26 by Zorila, Catalin, Boeddeker, Christoph, Doddipatla, Rama +1 · 1 citation
    #68T10 #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  4. Multiple-hypothesis CTC-based semi-supervised adaptation of end-to-end speech recognition
    2021/03/29 by Do, Cong-Thanh, Doddipatla, Rama, Hain, Thomas · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  5. Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis
    2024/07/04 by Cong-Thanh Do, Shuhei Imai, Do, Cong-Thanh +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis
  6. Prompting Whisper for QA-driven Zero-shot End-to-end Spoken Language Understanding
    2024/06/21 by Mohan Li, Li, Mohan, Simon Keizer +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Transformer-based Streaming ASR with Cumulative Attention
    2022/03/11 by Li, Mohan, Zhang, Shucong, Zorila, Catalin +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. WHISMA: A Speech-LLM to Perform Zero-shot Spoken Language Understanding
    2024/08/29 by Mohan Li, Li, Mohan, Cong-Thanh Do +9 · 2 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  9. Adversarial learning of neural user simulators for dialogue policy optimisation
    2023/06/01 by Keizer, Simon, Dockes, Caroline, Braunschweiler, Norbert +2 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  10. Non-autoregressive End-to-end Approaches for Joint Automatic Speech Recognition and Spoken Language Understanding
    2023/04/21 by Li, Mohan, Doddipatla, Rama · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. Evaluating Large Language Models for Document-grounded Response Generation in Information-Seeking Dialogues
    2023/09/21 by Braunschweiler, Norbert, Doddipatla, Rama, Keizer, Simon +1 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #I.2.7
  12. Frame-wise and overlap-robust speaker embeddings for meeting diarization
    2023/06/01 by Cord-Landwehr, Tobias, Boeddeker, Christoph, Zorilă, Cătălin +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  13. Monaural source separation: From anechoic to reverberant environments
    2021/11/15 by Cord-Landwehr, Tobias, Boeddeker, Christoph, von Neumann, Thilo +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  14. Transformer-based Online Speech Recognition with Decoder-end Adaptive Computation Steps
    2020/11/27 by Mohan Li, Cătălin Zorilă, Li, Mohan +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering