vix.ing · top · new · best · stats · spec

Serdyuk, Dmitriy

  1. Deep Complex Networks
    2017/05/27 by Trabelsi, Chiheb, Bilaniuk, Olexa, Zhang, Ying +7 · 19 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
  2. Attention-Based Models for Speech Recognition
    2015/06/24 by Chorowski, Jan, Bahdanau, Dzmitry, Serdyuk, Dmitriy +2 · 17 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
  3. Accounting for Variance in Machine Learning Benchmarks
    2021/03/01 by Xavier Bouthillier, Bouthillier, Xavier, Pierre Delaunay +31 · 2 voices · 15 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Machine Learning and Data Classification #Domain Adaptation and Few-Shot Learning
  4. End-to-End Attention-based Large Vocabulary Speech Recognition
    2015/08/18 by Bahdanau, Dzmitry, Chorowski, Jan, Serdyuk, Dmitriy +2 · 3 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
  5. Towards end-to-end spoken language understanding
    2018/02/23 by Serdyuk, Dmitriy, Wang, Yongqiang, Fuegen, Christian +3 · 2 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  6. Transformer-Based Video Front-Ends for Audio-Visual Speech Recognition for Single and Multi-Person Video
    2022/01/25 by Dmitriy Serdyuk, Otavio Braga, Serdyuk, Dmitriy +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Theano: A Python framework for fast computation of mathematical expressions
    2016/05/09 by The Theano Development Team, Al-Rfou, Rami, Alain, Guillaume +110 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mathematical Software (cs.MS) #Symbolic Computation (cs.SC)
  8. Invariant Representations for Noisy Speech Recognition
    2016/11/27 by Serdyuk, Dmitriy, Audhkhasi, Kartik, Brakel, Philémon +3 · 1 citation
    #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD)
  9. Unsupervised adversarial domain adaptation for acoustic scene classification
    2018/08/17 by Gharib, Shayan, Drossos, Konstantinos, Çakir, Emre +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  10. Audio-Visual Speech Recognition is Worth 32×32×8 Voxels
    2021/09/20 by Dmitriy Serdyuk, Otavio Braga, Serdyuk, Dmitriy +3 · 1 citation
    Computer Science · #Advanced Data Compression Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Speech and Audio Processing