Elsner, Micha
- Do Audio LLMs Really LISTEN, or Just Transcribe? Measuring Lexical vs. Acoustic Emotion Cues Reliance
2025/10/12 by Jingyi Chen, Zhimeng Guo, Chen, Jingyi +9 · 3 voices · 3 citations
Computer Science · #cs.CL #cs.AI
- DLPO: Diffusion Model Loss-Guided Reinforcement Learning for Fine-Tuning Text-to-Speech Diffusion Models
2024/05/23 by Chen, Jingyi, Byun, Ju-Seung, Elsner, Micha +1 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Exploring How Generative Adversarial Networks Learn Phonological Representations
2023/05/21 by Jingyi Chen, Micha Elsner, Chen, Jingyi +1 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Shortcomings of LLMs for Low-Resource Translation: Retrieval and Understanding are Both the Problem
2024/06/21 by Sara Court, Micha Elsner, Court, Sara +1 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Library Science and Information Systems #Machine Learning (cs.LG) #Natural Language Processing Techniques
- Analogy in Contact: Modeling Maltese Plural Inflection
2023/05/20 by Court, Sara, Sims, Andrea D., Elsner, Micha · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences