vix.ing · top · new · best · stats · spec

Damien Vincent

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1374 citations
    #cs.CL #cs.AI
  2. Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
    2024/03/08 by Gemini Robotics Team, Gemini Team, Petko Georgiev +2277 · 4 voices · 559 citations
    Computer Science · #Semantic Web and Ontologies
  3. AudioPaLM: A Large Language Model That Can Speak and Listen
    2023/06/22 by Paul K. Rubenstein, Rubenstein, Paul K., Chulayuth Asawaroengchai +57 · 1 voice · 52 citations
    #cs.CL #cs.AI #cs.SD #eess.AS #stat.ML
  4. Full Resolution Image Compression with Recurrent Neural Networks
    2016/08/18 by George Toderici, Toderici, George, Damien Vincent +11 · 2 voices · 4 citations
    #cs.CV
  5. AudioLM: a Language Modeling Approach to Audio Generation
    2022/09/07 by Zalán Borsos, Raphaël Marinier, Borsos, Zalán +18 · 94 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  6. Variable Rate Image Compression with Recurrent Neural Networks
    2015/11/19 by George Toderici, Toderici, George, Sean M. O’Malley +13 · 17 citations
    Computer Science · #Advanced Data Compression Techniques #Advanced Image Processing Techniques #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE)
  7. Speak, Read and Prompt: High-Fidelity Text-to-Speech with Minimal Supervision
    2023/02/07 by Eugene Kharitonov, Damien Vincent, Kharitonov, Eugene +15 · 27 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  8. SoundStorm: Efficient Parallel Audio Generation
    2023/05/16 by Zalán Borsos, Matt Sharifi, Borsos, Zalán +9 · 16 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Episodic Curiosity through Reachability
    2018/10/04 by Nikolay Savinov, Anton Raichuk, Savinov, Nikolay +11 · 16 citations
    Computer Science · #Reinforcement Learning in Robotics #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning
  10. RLDS: an Ecosystem to Generate, Share and Use Datasets in Reinforcement Learning
    2021/11/04 by Sabela Ramos, Sertan Girgin, Ramos, Sabela +21 · 9 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Scientific Computing and Data Management #Data Stream Mining Techniques
  11. Competitive Training of Mixtures of Independent Deep Generative Models
    2018/04/30 by Francesco Locatello, Damien Vincent, Locatello, Francesco +9 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification