vix.ing · top · new · best · stats · spec

Subramanian, Aswin Shanmugam

  1. CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
    2020/04/20 by Shinji Watanabe, Watanabe, Shinji, Michael Mandel +39 · 22 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Deep Learning based Multi-Source Localization with Source Splitting and its Effectiveness in Multi-Talker Speech Recognition
    2021/02/16 by Subramanian, Aswin Shanmugam, Weng, Chao, Watanabe, Shinji +2 · 4 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. Building state-of-the-art distant speech recognition using the CHiME-4 challenge with a setup of speech enhancement baseline
    2018/03/27 by Chen, Szu-Jui, Subramanian, Aswin Shanmugam, Xu, Hainan +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. The 2020 ESPnet update: new features, broadened applications, performance improvements, and future plans
    2020/12/23 by Shinji Watanabe, Florian Boyer, Watanabe, Shinji +27 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Tackling the Cocktail Fork Problem for Separation and Transcription of Real-World Soundtracks
    2022/12/14 by Petermann, Darius, Wichern, Gordon, Subramanian, Aswin Shanmugam +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  6. An Investigation of End-to-End Multichannel Speech Recognition for\n Reverberant and Mismatch Conditions
    2019/04/18 by Aswin Shanmugam Subramanian, Subramanian, Aswin Shanmugam, Xiaofei Wang +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering