vix.ing · top · new · best · stats · spec

Défossez, Alexandre

  1. High-Fidelity Simultaneous Speech-To-Speech Translation
    2025/02/05 by Tom Labiausse, Laurent Mazaré, Labiausse, Tom +10 · 13 voices · 7 citations
    Computer Science · #Speech Recognition and Synthesis #cs.CL #cs.SD #eess.AS
  2. High Fidelity Neural Audio Compression
    2022/10/24 by Alexandre Défossez, Défossez, Alexandre, Jade Copet +5 · 1 voice · 191 citations
    Computer Science · Engineering · Mathematics · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Signal Denoising Methods #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #cs.AI #cs.SD #eess.AS #electronic engineering #information engineering #stat.ML
  3. Code Llama: Open Foundation Models for Code
    2023/08/24 by Baptiste Rozière, Jonas Gehring, Rozière, Baptiste +48 · 341 citations
    Computer Science · #Model-Driven Software Engineering Techniques #Advanced Database Systems and Queries #Semantic Web and Ontologies
  4. Simple and Controllable Music Generation
    2023/06/08 by Jade Copet, Felix Kreuk, Copet, Jade +13 · 2 voices · 114 citations
    Computer Science · #Music Technology and Sound Studies #Music and Audio Processing #Speech and Audio Processing #cs.AI #cs.LG #cs.SD #eess.AS
  5. Moshi: a speech-text foundation model for real-time dialogue
    2024/09/17 by Alexandre Défossez, Défossez, Alexandre, Laurent Mazaré +13 · 156 citations
    Computer Science · #Speech and dialogue systems #Natural Language Processing Techniques #Multi-Agent Systems and Negotiation
  6. AudioGen: Textually Guided Audio Generation
    2022/09/30 by Felix Kreuk, Gabriel Synnaeve, Kreuk, Felix +15 · 52 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Hybrid Transformers for Music Source Separation
    2022/11/15 by Simon Rouard, Rouard, Simon, Francisco Massa +3 · 37 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Demucs: Deep Extractor for Music Sources with extra unlabeled data remixed
    2019/09/03 by Alexandre Défossez, Nicolas Usunier, Défossez, Alexandre +5 · 15 citations
    Computer Science · #Speech and Audio Processing #Music and Audio Processing #Speech Recognition and Synthesis
  9. A Simple Convergence Proof of Adam and Adagrad
    2020/03/05 by Alexandre Défossez, Léon Bottou, Défossez, Alexandre +5 · 13 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sparse and Compressive Sensing Techniques #Stochastic Gradient Optimization Techniques
  10. Proactive Detection of Voice Cloning with Localized Watermarking
    2024/01/30 by Robin San Roman, Pierre Fernandez, Roman, Robin San +9 · 25 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #Digital Media Forensic Detection #FOS: Computer and information sciences #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing
  11. Music Source Separation in the Waveform Domain
    2019/11/27 by Alexandre Défossez, Nicolas Usunier, Défossez, Alexandre +5 · 17 citations
    Computer Science · #Speech and Audio Processing #Music and Audio Processing #Speech Recognition and Synthesis
  12. Hybrid Spectrogram and Waveform Source Separation
    2021/11/05 by Alexandre Défossez, Défossez, Alexandre · 14 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  13. Masked Audio Generation using a Single Non-Autoregressive Transformer
    2024/01/09 by Alon Ziv, Ziv, Alon, Itai Gat +15 · 15 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  14. Differentiable Model Compression via Pseudo Quantization Noise
    2021/04/20 by Alexandre Défossez, Yossi Adi, Défossez, Alexandre +3 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  15. An Independence-promoting Loss for Music Generation with Language Models
    2024/06/04 by Jean-Marie Lemercier, Lemercier, Jean-Marie, Simon Rouard +6 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  16. Audio Conditioning for Music Generation via Discrete Bottleneck Features
    2024/07/17 by Simon Rouard, Rouard, Simon, Yossi Adi +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  17. Vision-Speech Models: Teaching Speech Models to Converse about Images
    2025/03/19 by Amélie Royer, Moritz Böhle, Royer, Amélie +11 · 2 citations
    Earth and Planetary Sciences · Social Sciences · #3D Surveying and Cultural Heritage #Computer Vision and Pattern Recognition (cs.CV) #Educational Tools and Methods #FOS: Computer and information sciences
  18. Aligning Spoken Dialogue Models from User Interactions
    2025/06/26 by Anqi Wu, Laurent Mazaré, Wu, Anne +5 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  19. Streaming Sequence-to-Sequence Learning with Delayed Streams Modeling
    2025/09/10 by Neil Zeghidour, Zeghidour, Neil, Eugene Kharitonov +15 · 5 citations
    Computer Science · Engineering · #Computation and Language (cs.CL) #Data Stream Mining Techniques #FOS: Computer and information sciences #Fault Detection and Control Systems #Neural Networks and Applications