vix.ing · top · new · best · stats · spec

Pellegrini, Thomas

  1. CoNeTTE: An efficient Audio Captioning system leveraging multiple datasets with Task Embedding
    2023/09/01 by Étienne Labbé, Labbé, Étienne, Thomas Pellegrini +3 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Video Analysis and Summarization #electronic engineering #information engineering
  2. Low-activity supervised convolutional spiking neural networks applied to speech commands recognition
    2020/11/13 by Pellegrini, Thomas, Zimmer, Romain, Masquelier, Timothée · 3 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  3. Adapting a ConvNeXt model to audio classification on AudioSet
    2023/06/01 by Thomas Pellegrini, Pellegrini, Thomas, Ismail Khalfaoui-Hassani +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
    2025/06/25 by Tuncay, Ludovic, Labbé, Etienne, Benetos, Emmanouil +1 · 4 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Signal Processing (eess.SP) #Sound (cs.SD) #electronic engineering #information engineering
  5. Evaluation of post-processing algorithms for polyphonic sound event\n detection
    2019/06/17 by Léo Cances, Cances, Leo, Patrice Guyot +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  6. Dilated Convolution with Learnable Spacings: beyond bilinear interpolation
    2023/06/01 by Ismail Khalfaoui-Hassani, Thomas Pellegrini, Khalfaoui-Hassani, Ismail +3 · 2 citations
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  7. End-to-end acoustic modelling for phone recognition of young readers
    2021/03/04 by Gelin, Lucile, Daniel, Morgane, Pinquier, Julien +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Is my automatic audio captioning system so bad? spider-max: a metric to consider several caption candidates
    2022/11/14 by Labbé, Etienne, Pellegrini, Thomas, Pinquier, Julien · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering