Serdyuk, Dmitriy
- Deep Complex Networks
2017/05/27 by Trabelsi, Chiheb, Bilaniuk, Olexa, Zhang, Ying +7 · 19 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
- Attention-Based Models for Speech Recognition
2015/06/24 by Chorowski, Jan, Bahdanau, Dzmitry, Serdyuk, Dmitriy +2 · 17 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
- Accounting for Variance in Machine Learning Benchmarks
2021/03/01 by Xavier Bouthillier, Bouthillier, Xavier, Pierre Delaunay +31 · 2 voices · 15 citations
Computer Science · #Adversarial Robustness in Machine Learning #Machine Learning and Data Classification #Domain Adaptation and Few-Shot Learning
- End-to-End Attention-based Large Vocabulary Speech Recognition
2015/08/18 by Bahdanau, Dzmitry, Chorowski, Jan, Serdyuk, Dmitriy +2 · 3 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
- Towards end-to-end spoken language understanding
2018/02/23 by Serdyuk, Dmitriy, Wang, Yongqiang, Fuegen, Christian +3 · 2 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Transformer-Based Video Front-Ends for Audio-Visual Speech Recognition for Single and Multi-Person Video
2022/01/25 by Dmitriy Serdyuk, Otavio Braga, Serdyuk, Dmitriy +3 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Theano: A Python framework for fast computation of mathematical expressions
2016/05/09 by The Theano Development Team, Al-Rfou, Rami, Alain, Guillaume +110 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Mathematical Software (cs.MS) #Symbolic Computation (cs.SC)
- Invariant Representations for Noisy Speech Recognition
2016/11/27 by Serdyuk, Dmitriy, Audhkhasi, Kartik, Brakel, Philémon +3 · 1 citation
#Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD)
- Unsupervised adversarial domain adaptation for acoustic scene classification
2018/08/17 by Gharib, Shayan, Drossos, Konstantinos, Çakir, Emre +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Audio-Visual Speech Recognition is Worth 32×32×8 Voxels
2021/09/20 by Dmitriy Serdyuk, Otavio Braga, Serdyuk, Dmitriy +3 · 1 citation
Computer Science · #Advanced Data Compression Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Speech and Audio Processing