Pascual, Santiago
- SEGAN: Speech Enhancement Generative Adversarial Network
2017/03/28 by Pascual, Santiago, Bonafonte, Antonio, Serrà, Joan · 23 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Sound (cs.SD)
- V2A-Mapper: A Lightweight Solution for Vision-to-Audio Generation by Connecting Foundation Models
2023/08/18 by Heng Wang, Jianbo Ma, Wang, Heng +7 · 22 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Universal Speech Enhancement with Score-based Diffusion
2022/06/07 by Serrà, Joan, Pascual, Santiago, Pons, Jordi +2 · 15 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Automatic multitrack mixing with a differentiable mixing console of neural audio effects
2020/10/20 by Steinmetz, Christian J., Pons, Jordi, Pascual, Santiago +1 · 10 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Upsampling artifacts in neural audio synthesis
2020/10/27 by Pons, Jordi, Pascual, Santiago, Cengarle, Giulio +1 · 6 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Blow: a single-scale hyperconditioned flow for non-parallel raw-audio\n voice conversion
2019/06/03 by Joan Serrà, Santiago Pascual, Serrà, Joan +3 · 7 citations
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
- Masked Generative Video-to-Audio Transformers with Enhanced Synchronicity
2024/07/15 by Santiago Pascual, Pascual, Santiago, Chunghsin Yeh +5 · 10 citations
Computer Science · Engineering · #Advanced Optical Imaging Technologies #Advanced Vision and Imaging #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Sound (cs.SD) #electronic engineering #information engineering
- On loss functions and evaluation metrics for music source separation
2022/02/16 by Enric Gusó, Gusó, Enric, Jordi Pons +5 · 5 citations
Computer Science · Engineering · #Speech and Audio Processing #Advanced Adaptive Filtering Techniques #Acoustic Wave Phenomena Research
- Full-band General Audio Synthesis with Score-based Diffusion
2022/10/26 by Santiago Pascual, Pascual, Santiago, Gautam Bhattacharya +7 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Multi-task self-supervised learning for Robust Speech Recognition
2020/01/25 by Ravanelli, Mirco, Zhong, Jianyuan, Pascual, Santiago +4 · 3 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- SESQA: semi-supervised learning for speech quality assessment
2020/10/01 by Serrà, Joan, Pons, Jordi, Pascual, Santiago · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Adversarial Permutation Invariant Training for Universal Sound Separation
2022/10/21 by Emilian Postolache, Jordi Pons, Postolache, Emilian +5 · 2 citations
Computer Science · Engineering · #Acoustic Wave Phenomena Research #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- GASS: Generalizing Audio Source Separation with Large-scale Data
2023/09/29 by Pons, Jordi, Liu, Xiaoyu, Pascual, Santiago +1 · 2 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Signal Processing (eess.SP) #Sound (cs.SD) #electronic engineering #information engineering
- Towards a universal neural network encoder for time series
2018/05/10 by Serrà, Joan, Pascual, Santiago, Karatzoglou, Alexandros · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
- Wav2Pix: Speech-conditioned Face Generation using Generative Adversarial Networks
2019/03/25 by Duarte, Amanda, Roldan, Francisco, Tubau, Miquel +7 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM)
- Towards Generalized Speech Enhancement with Generative Adversarial Networks
2019/04/06 by Pascual, Santiago, Serrà, Joan, Bonafonte, Antonio · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Adversarial Auto-Encoding for Packet Loss Concealment
2021/07/07 by Pascual, Santiago, Serrà, Joan, Pons, Jordi · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Mono-to-stereo through parametric stereo generation
2023/06/26 by Serrà, Joan, Scaini, Davide, Pascual, Santiago +4 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Sequential Contrastive Audio-Visual Learning
2024/07/08 by Ioannis Tsiamas, Santiago Pascual, Tsiamas, Ioannis +5 · 1 voice · 1 citation
#cs.SD #cs.CV #cs.LG #cs.MM #eess.AS
- Joint Semantic Knowledge Distillation and Masked Acoustic Modeling for Full-band Speech Restoration with Improved Intelligibility
2024/09/14 by Xiaoyu Liu, Xu Li, Liu, Xiaoyu +5 · 1 citation
Computer Science · Medicine · #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders