Anton Ragni
- MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
2023/05/31 by Yizhi Li, Ruibin Yuan, Li, Yizhi +35 · 76 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- MARBLE: Music Audio Representation Benchmark for Universal Evaluation
2023/06/18 by Ruibin Yuan, Yinghao Ma, Yuan, Ruibin +46 · 15 citations
Computer Science · Arts and Humanities · #Music and Audio Processing #Diverse Musicological Studies #Music Technology and Sound Studies
- MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning
2022/12/05 by Yizhi Li, Li, Yizhi, Ruibin Yuan +25 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Foundation Models for Music: A Survey
2024/08/26 by Yinghao Ma, Anders Øland, Ma, Yinghao +81 · 13 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Sound (cs.SD) #electronic engineering #information engineering
- Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
2024/06/12 by Wing-Zin Leung, Leung, Wing-Zin, Mattias Cross +5 · 10 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Non-Intrusive Speech Intelligibility Prediction for Hearing-Impaired Users using Intermediate ASR Features and Human Memory Models
2024/01/24 by Rhiannon Mogridge, Mogridge, Rhiannon, George Close +11 · 8 citations
Computer Science · Health Professions · Neuroscience · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Noise Effects and Management #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Approximate Fixed-Points in Recurrent Neural Networks
2021/06/04 by Zheng‐Xiong Wang, Anton Ragni, Wang, Zhengxiong +1 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Neural Networks and Applications #Topic Modeling #electronic engineering #information engineering
- On the Effectiveness of Speech Self-supervised Learning for Music
2023/07/11 by Yinghao Ma, Ma, Yinghao, Ruibin Yuan +27 · 1 citation
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Music Technology and Sound Studies
- HERB: Measuring Hierarchical Regional Bias in Pre-trained Language Models
2022/11/05 by Yizhi Li, Li, Yizhi, Ge Zhang +11 · 1 citation
Social Sciences · #Computation and Language (cs.CL) #Ethics and Social Impacts of AI #FOS: Computer and information sciences