Eugene Kharitonov
- AudioPaLM: A Large Language Model That Can Speak and Listen
2023/06/22 by Paul K. Rubenstein, Chulayuth Asawaroengchai, Rubenstein, Paul K. +57 · 1 voice · 47 citations
#cs.CL #cs.AI #cs.SD #eess.AS #stat.ML
- Gemma 3 Technical Report
2025/03/25 by Aishwarya Kamath, Gemma Team, Johan Ferret +418 · 3 voices · 464 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.AI #cs.CL
- AudioLM: a Language Modeling Approach to Audio Generation
2022/09/07 by Zalán Borsos, Raphaël Marinier, Borsos, Zalán +18 · 91 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Speech Resynthesis from Discrete Disentangled Self-Supervised Representations
2021/04/01 by Adam Polyak, Yossi Adi, Polyak, Adam +13 · 20 citations
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
2023/02/07 by Eugene Kharitonov, Kharitonov, Eugene, Damien Vincent +15 · 26 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
- SoundStorm: Efficient Parallel Audio Generation
2023/05/16 by Zalán Borsos, Borsos, Zalán, Matt Sharifi +9 · 15 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Text-Free Prosody-Aware Generative Spoken Language Modeling
2021/09/07 by Eugene Kharitonov, Ann Lee, Kharitonov, Eugene +19 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Anti-efficient encoding in emergent communication
2019/05/29 by Rahma Chaabouni, Chaabouni, Rahma, Eugene Kharitonov +5 · 5 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Cellular Automata and Applications #Computation and Language (cs.CL) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Language and cultural evolution #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
- Textless Speech Emotion Conversion using Discrete and Decomposed Representations
2021/11/14 by Felix Kreuk, Adam Polyak, Kreuk, Felix +16 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Emergent Language Generalization and Acquisition Speed are not tied to\n Compositionality
2020/04/07 by Eugene Kharitonov, Kharitonov, Eugene, Marco Baroni +1 · 1 citation
Computer Science · Engineering · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Machine Learning (cs.LG) #Natural Language Processing Techniques #Neural Networks and Applications
- What they do when in doubt: a study of inductive biases in seq2seq learners
2020/06/26 by Eugene Kharitonov, Rahma Chaabouni, Kharitonov, Eugene +1 · 1 citation
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Information Theory (cs.IT) #Machine Learning and Algorithms
- textless-lib: a Library for Textless Spoken Language Processing
2022/02/15 by Eugene Kharitonov, Jade Copet, Kharitonov, Eugene +19 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- MAD Speech: Measures of Acoustic Diversity of Speech
2024/04/16 by Matthieu Futeral, Futeral, Matthieu, Andrea Agostinelli +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering
- Streaming Sequence-to-Sequence Learning with Delayed Streams Modeling
2025/09/10 by Neil Zeghidour, Eugene Kharitonov, Zeghidour, Neil +15 · 5 citations
Computer Science · Engineering · #Computation and Language (cs.CL) #Data Stream Mining Techniques #FOS: Computer and information sciences #Fault Detection and Control Systems #Neural Networks and Applications