vix.ing · top · new · best · stats · spec

Jesse Engel

  1. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin
    2015/12/08 by Dario Amodei, Amodei, Dario, Rishita Anubhai +66 · 1 voice · 119 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL
  2. MuChoMusic dataset
    2023/01/26 by Andrea Agostinelli, Timo I. Denk, Agostinelli, Andrea +23 · 1 voice · 89 citations
    Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech Recognition and Synthesis #cs.LG #cs.SD #eess.AS
  3. Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
    2022/06/09 by Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448 · 3 voices · 175 citations
    #cs.CL #cs.AI #cs.CY #cs.LG #stat.ML
  4. Neural Audio Synthesis of Musical Notes with WaveNet Autoencoders
    2017/04/05 by Jesse Engel, Cinjon Resnick, Engel, Jesse +11 · 54 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing
  5. Enabling Factorized Piano Music Modeling and Generation with the MAESTRO\n Dataset
    2018/10/29 by Curtis Hawthorne, Hawthorne, Curtis, Andriy Stasyuk +15 · 50 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  6. DDSP: Differentiable Digital Signal Processing
    2020/01/14 by Jesse Engel, Engel, Jesse, Lamtharn Hantrakul +6 · 1 voice · 38 citations
    Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #cs.LG #cs.SD #eess.AS #eess.SP #stat.ML
  7. Onsets and Frames: Dual-Objective Piano Transcription
    2017/10/30 by Curtis Hawthorne, Erich Elsen, Hawthorne, Curtis +15 · 17 citations
    Computer Science · Neuroscience · #Music and Audio Processing #Music Technology and Sound Studies #Neuroscience and Music Perception
  8. GANSynth: Adversarial Neural Audio Synthesis
    2019/02/23 by Jesse Engel, Kumar Krishna Agrawal, Engel, Jesse +9 · 24 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing #Digital Media Forensic Detection
  9. HEAR: Holistic Evaluation of Audio Representations
    2022/03/06 by Joseph Turian, Jordie Shier, Turian, Joseph +43 · 22 citations
    Computer Science · Neuroscience · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  10. Noise2Music: Text-conditioned Music Generation with Diffusion Models
    2023/02/08 by Qingqing Huang, Daniel Park, Huang, Qingqing +25 · 19 citations
    Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech Recognition and Synthesis
  11. MT3: Multi-Task Multitrack Music Transcription
    2021/11/04 by Josh Gardner, Ian Simon, Gardner, Josh +7 · 11 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  12. Symbolic Music Generation with Diffusion Models
    2021/03/30 by Gautam Mittal, Jesse Engel, Mittal, Gautam +5 · 12 citations
    Computer Science · Physics and Astronomy · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  13. Sequence-to-Sequence Piano Transcription with Transformers
    2021/07/19 by Curtis Hawthorne, Ian Simon, Hawthorne, Curtis +7 · 10 citations
    Computer Science · Arts and Humanities · #Music and Audio Processing #Diverse Musicological Studies #Music Technology and Sound Studies
  14. MIDI-DDSP: Detailed Control of Musical Performance via Hierarchical Modeling
    2021/12/17 by Yusong Wu, Ethan Manilow, Wu, Yusong +15 · 9 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
  15. Latent Constraints: Learning to Generate Conditionally from Unconditional Generative Models
    2017/11/15 by Jesse Engel, Matthew Hoffman, Engel, Jesse +5 · 7 citations
    Computer Science · Physics and Astronomy · #Computer Graphics and Visualization Techniques #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Neural and Evolutionary Computing (cs.NE)
  16. SingSong: Generating musical accompaniments from singing
    2023/01/30 by Chris Donahue, Donahue, Chris, Antoine Caillon +19 · 11 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  17. Improving Perceptual Quality of Drum Transcription with the Expanded Groove MIDI Dataset
    2020/04/01 by Lee Callender, Callender, Lee, Curtis Hawthorne +3 · 6 citations
    Arts and Humanities · Computer Science · #Diverse Musicological Studies #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD)
  18. Encoding Musical Style with Transformer Autoencoders
    2019/12/10 by Kristy Choi, Curtis Hawthorne, Choi, Kristy +7 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  19. General-purpose, long-context autoregressive modeling with Perceiver AR
    2022/02/15 by Curtis Hawthorne, Andrew Jaegle, Hawthorne, Curtis +27 · 4 citations
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  20. The Chamber Ensemble Generator: Limitless High-Quality MIR Data via Generative Modeling
    2022/09/28 by Yusong Wu, Wu, Yusong, Josh Gardner +9 · 3 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering
  21. Latent Translation: Crossing Modalities by Bridging Generative Models
    2019/02/21 by Yingtao Tian, Tian, Yingtao, Jesse Engel +1 · 3 citations
    Computer Science · #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis
  22. Learning a Latent Space of Multitrack Measures
    2018/06/01 by Ian Simon, Adam P. Roberts, Simon, Ian +9 · 1 citation
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music Technology and Sound Studies #Music and Audio Processing #Neuroscience and Music Perception #Sound (cs.SD) #electronic engineering #information engineering