Huybrechts, Goeric
- The Amazon Nova Family of Models: Technical Report and Model Card
2025/03/17 by Amazon AGI, AGI, Amazon, Langford, Aaron +867 · 26 citations
Computer Science · Engineering · #3D Modeling in Geospatial Applications #Artificial Intelligence (cs.AI) #BIM and Construction Integration #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model-Driven Software Engineering Techniques
- Dynamic Chunk Convolution for Unified Streaming and Non-Streaming Conformer ASR
2023/04/18 by Xilai Li, Goeric Huybrechts, Li, Xilai +7 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Wanda++: Pruning Large Language Models via Regional Gradients
2025/03/06 by Yifan Yang, Yang, Yifan, Kai Zhen +24 · 11 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
2024/10/26 by Sullam Jeoung, Jeoung, Sullam, Goeric Huybrechts +7 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics
- EmoCat: Language-agnostic Emotional Voice Conversion
2021/01/14 by Schnell, Bastian, Huybrechts, Goeric, Perz, Bartek +2 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Low-resource expressive text-to-speech using data augmentation
2020/11/11 by Huybrechts, Goeric, Merritt, Thomas, Comini, Giulia +3 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module
2022/02/16 by Adam Gabryś, Goeric Huybrechts, Gabryś, Adam +15 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models
2024/05/14 by Raghuveer Peri, Sai Muralidhar Jayanthi, Peri, Raghuveer +25 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Zero-resource Speech Translation and Recognition with LLMs
2024/12/24 by Karel Mundnich, Xing Niu, Mundnich, Karel +23 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #electronic engineering #information engineering
- Non-Autoregressive TTS with Explicit Duration Modelling for Low-Resource\n Highly Expressive Speech
2021/06/24 by Raahil Shah, Kamil Pokora, Shah, Raahil +13 · 1 citation
Computer Science · Physics and Astronomy · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling
- Cross-speaker style transfer for text-to-speech using data augmentation
2022/02/10 by Manuel Sam Ribeiro, Julian Roth, Ribeiro, Manuel Sam +9 · 1 citation
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation
2022/07/29 by Giulia Comini, Goeric Huybrechts, Comini, Giulia +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering