vix.ing · top · new · best · stats · spec

Huybrechts, Goeric

  1. The Amazon Nova Family of Models: Technical Report and Model Card
    2025/03/17 by Amazon AGI, AGI, Amazon, Langford, Aaron +867 · 26 citations
    Computer Science · Engineering · #3D Modeling in Geospatial Applications #Artificial Intelligence (cs.AI) #BIM and Construction Integration #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model-Driven Software Engineering Techniques
  2. Dynamic Chunk Convolution for Unified Streaming and Non-Streaming Conformer ASR
    2023/04/18 by Xilai Li, Goeric Huybrechts, Li, Xilai +7 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Wanda++: Pruning Large Language Models via Regional Gradients
    2025/03/06 by Yifan Yang, Yang, Yifan, Kai Zhen +24 · 11 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  4. Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
    2024/10/26 by Sullam Jeoung, Jeoung, Sullam, Goeric Huybrechts +7 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics
  5. EmoCat: Language-agnostic Emotional Voice Conversion
    2021/01/14 by Schnell, Bastian, Huybrechts, Goeric, Perz, Bartek +2 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  6. Low-resource expressive text-to-speech using data augmentation
    2020/11/11 by Huybrechts, Goeric, Merritt, Thomas, Comini, Giulia +3 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module
    2022/02/16 by Adam Gabryś, Goeric Huybrechts, Gabryś, Adam +15 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  8. SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models
    2024/05/14 by Raghuveer Peri, Sai Muralidhar Jayanthi, Peri, Raghuveer +25 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  9. Zero-resource Speech Translation and Recognition with LLMs
    2024/12/24 by Karel Mundnich, Xing Niu, Mundnich, Karel +23 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #electronic engineering #information engineering
  10. Non-Autoregressive TTS with Explicit Duration Modelling for Low-Resource\n Highly Expressive Speech
    2021/06/24 by Raahil Shah, Kamil Pokora, Shah, Raahil +13 · 1 citation
    Computer Science · Physics and Astronomy · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling
  11. Cross-speaker style transfer for text-to-speech using data augmentation
    2022/02/10 by Manuel Sam Ribeiro, Julian Roth, Ribeiro, Manuel Sam +9 · 1 citation
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  12. Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation
    2022/07/29 by Giulia Comini, Goeric Huybrechts, Comini, Giulia +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering