Jesse Engel
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
2015/12/08 by Dario Amodei, Amodei, Dario, Rishita Anubhai +66 · 1 voice · 119 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL
- MuChoMusic dataset
2023/01/26 by Andrea Agostinelli, Timo I. Denk, Agostinelli, Andrea +23 · 1 voice · 89 citations
Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech Recognition and Synthesis #cs.LG #cs.SD #eess.AS
- Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
2022/06/09 by Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448 · 3 voices · 175 citations
#cs.CL #cs.AI #cs.CY #cs.LG #stat.ML
- Neural Audio Synthesis of Musical Notes with WaveNet Autoencoders
2017/04/05 by Jesse Engel, Cinjon Resnick, Engel, Jesse +11 · 54 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing
- Enabling Factorized Piano Music Modeling and Generation with the MAESTRO\n Dataset
2018/10/29 by Curtis Hawthorne, Hawthorne, Curtis, Andriy Stasyuk +15 · 50 citations
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- DDSP: Differentiable Digital Signal Processing
2020/01/14 by Jesse Engel, Engel, Jesse, Lamtharn Hantrakul +6 · 1 voice · 38 citations
Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #cs.LG #cs.SD #eess.AS #eess.SP #stat.ML
- Onsets and Frames: Dual-Objective Piano Transcription
2017/10/30 by Curtis Hawthorne, Erich Elsen, Hawthorne, Curtis +15 · 17 citations
Computer Science · Neuroscience · #Music and Audio Processing #Music Technology and Sound Studies #Neuroscience and Music Perception
- GANSynth: Adversarial Neural Audio Synthesis
2019/02/23 by Jesse Engel, Kumar Krishna Agrawal, Engel, Jesse +9 · 24 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing #Digital Media Forensic Detection
- HEAR: Holistic Evaluation of Audio Representations
2022/03/06 by Joseph Turian, Jordie Shier, Turian, Joseph +43 · 22 citations
Computer Science · Neuroscience · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Noise2Music: Text-conditioned Music Generation with Diffusion Models
2023/02/08 by Qingqing Huang, Daniel Park, Huang, Qingqing +25 · 19 citations
Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech Recognition and Synthesis
- MT3: Multi-Task Multitrack Music Transcription
2021/11/04 by Josh Gardner, Ian Simon, Gardner, Josh +7 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Symbolic Music Generation with Diffusion Models
2021/03/30 by Gautam Mittal, Jesse Engel, Mittal, Gautam +5 · 12 citations
Computer Science · Physics and Astronomy · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- Sequence-to-Sequence Piano Transcription with Transformers
2021/07/19 by Curtis Hawthorne, Ian Simon, Hawthorne, Curtis +7 · 10 citations
Computer Science · Arts and Humanities · #Music and Audio Processing #Diverse Musicological Studies #Music Technology and Sound Studies
- MIDI-DDSP: Detailed Control of Musical Performance via Hierarchical Modeling
2021/12/17 by Yusong Wu, Ethan Manilow, Wu, Yusong +15 · 9 citations
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
- Latent Constraints: Learning to Generate Conditionally from Unconditional Generative Models
2017/11/15 by Jesse Engel, Matthew Hoffman, Engel, Jesse +5 · 7 citations
Computer Science · Physics and Astronomy · #Computer Graphics and Visualization Techniques #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Neural and Evolutionary Computing (cs.NE)
- SingSong: Generating musical accompaniments from singing
2023/01/30 by Chris Donahue, Donahue, Chris, Antoine Caillon +19 · 11 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Improving Perceptual Quality of Drum Transcription with the Expanded Groove MIDI Dataset
2020/04/01 by Lee Callender, Callender, Lee, Curtis Hawthorne +3 · 6 citations
Arts and Humanities · Computer Science · #Diverse Musicological Studies #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD)
- Encoding Musical Style with Transformer Autoencoders
2019/12/10 by Kristy Choi, Curtis Hawthorne, Choi, Kristy +7 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- General-purpose, long-context autoregressive modeling with Perceiver AR
2022/02/15 by Curtis Hawthorne, Andrew Jaegle, Hawthorne, Curtis +27 · 4 citations
Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- The Chamber Ensemble Generator: Limitless High-Quality MIR Data via Generative Modeling
2022/09/28 by Yusong Wu, Wu, Yusong, Josh Gardner +9 · 3 citations
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
- Latent Translation: Crossing Modalities by Bridging Generative Models
2019/02/21 by Yingtao Tian, Tian, Yingtao, Jesse Engel +1 · 3 citations
Computer Science · #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis
- Learning a Latent Space of Multitrack Measures
2018/06/01 by Ian Simon, Adam P. Roberts, Simon, Ian +9 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering