vix.ing · top · new · best · stats · spec

Björn W. Schuller

  1. HEAR: Holistic Evaluation of Audio Representations
    2022/03/06 by Joseph Turian, Turian, Joseph, Jordie Shier +43 · 17 citations
    Computer Science · Neuroscience · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  2. The Age of Artificial Emotional Intelligence
    2018/09/01 by Dagmar Schuller, Björn W. Schuller, Bjorn W. Schuller · 16 citations
    Psychology · Neuroscience · #Emotional Intelligence and Performance #Emotion and Mood Recognition #Neuroscience, Education and Cognitive Function
  3. Recognising realistic emotions and affect in speech: State of the art and lessons learnt from the first challenge
    2011/02/07 by Björn W. Schuller, Björn Schuller, Anton Batliner +2 · 10 citations
    Computer Science · Psychology · #Emotion and Mood Recognition #Music and Audio Processing #Speech Recognition and Synthesis
  4. Affective Image Content Analysis: Two Decades Review and New Perspectives
    2021/06/30 by Sicheng Zhao, Xingxu Yao, Zhao, Sicheng +13 · 9 citations
    Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Emotion and Mood Recognition #FOS: Computer and information sciences #Image Retrieval and Classification Techniques #Multimedia (cs.MM) #Sentiment Analysis and Opinion Mining
  5. MER 2023: Multi-label Learning, Modality Robustness, and Semi-Supervised Learning
    2023/04/18 by Zheng Lian, Lian, Zheng, Haiyang Sun +27 · 11 citations
    Psychology · Computer Science · #Emotion and Mood Recognition #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies
  6. auDeep: Unsupervised Learning of Representations from Audio with Deep Recurrent Neural Networks
    2017/12/12 by Michael Freitag, Freitag, Michael, Shahin Amiriparian +8 · 8 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Music Technology and Sound Studies
  7. Convolutional RNN: an Enhanced Model for Extracting Features from Sequential Data
    2016/02/18 by Gil Keren, Keren, Gil, Björn W. Schuller +1 · 4 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
  8. Paralinguistics in speech and language—State-of-the-art and the challenge
    2012/03/08 by Björn W. Schuller, Björn Schuller, Stefan Steidl +5 · 4 citations
    Psychology · #Emotion and Mood Recognition #Multisensory perception and integration #Phonetics and Phonology Research
  9. Audio Barlow Twins: Self-Supervised Audio Representation Learning
    2022/09/28 by Jonah Anton, Harry Coppock, Anton, Jonah +5 · 5 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  10. MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition
    2024/04/26 by Zheng Lian, Haiyang Sun, Lian, Zheng +33 · 8 citations
    Psychology · #Emotion and Mood Recognition #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG)
  11. Self Supervised Adversarial Domain Adaptation for Cross-Corpus and Cross-Language Speech Emotion Recognition
    2022/04/19 by Siddique Latif, Latif, Siddique, Rajib Rana +7 · 4 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  12. Refashioning Emotion Recognition Modelling: The Advent of Generalised Large Models
    2023/08/21 by Zixing Zhang, Zhang, Zixing, Liyizhe Peng +9 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Sentiment Analysis and Opinion Mining #Topic Modeling
  13. PerCo (SD): Open Perceptual Compression
    2024/09/30 by Nikolai Körber, Körber, Nikolai, Eduard Kromer +9 · 6 citations
    Computer Science · #Advanced Data Compression Techniques #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  14. Deep Learning for Environmentally Robust Speech Recognition: An Overview of Recent Developments
    2017/05/30 by Zixing Zhang, Jürgen T. Geiger, Zhang, Zixing +9 · 2 citations
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  15. ExHuBERT: Enhancing HuBERT Through Block Extension and Fine-Tuning on 37 Emotion Datasets
    2024/06/11 by Shahin Amiriparian, Amiriparian, Shahin, Filip Packań +5 · 5 citations
    Psychology · #68T10 #Computation and Language (cs.CL) #Emotion and Mood Recognition #FOS: Computer and information sciences #I.2
  16. Multi-Task Semi-Supervised Adversarial Autoencoding for Speech Emotion\n Recognition
    2019/07/13 by Siddique Latif, Rajib Rana, Latif, Siddique +9 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  17. The Multimodal Sentiment Analysis in Car Reviews (MuSe-CaR) Dataset:\n Collection, Insights and Improvements
    2021/01/15 by Lukas Stappen, Alice Baird, Stappen, Lukas +5 · 2 citations
    Computer Science · Psychology · #Sentiment Analysis and Opinion Mining #Emotion and Mood Recognition #Advanced Text Analysis Techniques
  18. Towards Multimodal Prediction of Spontaneous Humour: A Novel Dataset and First Results
    2022/09/28 by Lukas Christ, Christ, Lukas, Shahin Amiriparian +9 · 2 citations
    Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Humor Studies and Applications #Machine Learning (cs.LG) #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
  19. Are 3D Face Shapes Expressive Enough for Recognising Continuous Emotions and Action Unit Intensities?
    2022/07/03 by Mani Kumar Tellamekala, Tellamekala, Mani Kumar, Ömer Sümer +9 · 2 citations
    Computer Science · Neuroscience · Psychology · #Computer Vision and Pattern Recognition (cs.CV) #Emotion and Mood Recognition #FOS: Computer and information sciences #Face Recognition and Perception #Face recognition and analysis #Human-Computer Interaction (cs.HC) #Multimedia (cs.MM)
  20. A large-scale and PCR-referenced vocal audio dataset for COVID-19
    2022/12/15 by Jobie Budd, Kieran Baker, Budd, Jobie +49 · 2 citations
    Medicine · #Audio and Speech Processing (eess.AS) #COVID-19 diagnosis using AI #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Phonocardiography and Auscultation Techniques #Respiratory and Cough-Related Research #Sound (cs.SD) #electronic engineering #information engineering
  21. Explainable Artificial Intelligence for Medical Applications: A Review
    2024/11/15 by Qiyang Sun, Alican Akman, Sun, Qiyang +3 · 6 citations
    Computer Science · Neuroscience · Health Professions · #Machine Learning in Healthcare #Brain Tumor Detection and Classification #Artificial Intelligence in Healthcare
  22. openXBOW - Introducing the Passau Open-Source Crossmodal Bag-of-Words Toolkit
    2016/05/22 by Maximilian Schmitt, Björn W. Schuller, Schmitt, Maximilian +1 · 1 citation
    Computer Science · #Text and Document Classification Technologies #Music and Audio Processing #Advanced Text Analysis Techniques
  23. Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers
    2023/07/12 by Siddique Latif, Latif, Siddique, Muhammad Usama +5 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  24. Audio-based AI classifiers show no evidence of improved COVID-19 screening over simple symptoms checkers
    2022/12/15 by Harry Coppock, The Alan Turing Institute, Coppock, Harry +49 · 2 citations
    Medicine · Computer Science · #COVID-19 diagnosis using AI #Music and Audio Processing #Phonocardiography and Auscultation Techniques
  25. Deep Representation Learning in Speech Processing: Challenges, Recent Advances, and Future Trends
    2020/01/02 by Siddique Latif, Rajib Rana, Latif, Siddique +9 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  26. Computational Emotion Analysis From Images: Recent Advances and Future Directions
    2021/03/19 by Sicheng Zhao, Zhao, Sicheng, Quanwei Huang +11 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human-Computer Interaction (cs.HC) #Image Retrieval and Classification Techniques #Multimedia (cs.MM) #Sentiment Analysis and Opinion Mining
  27. DeepSpectrumLite: A Power-Efficient Transfer Learning Framework for Embedded Speech and Audio Processing from Decentralised Data
    2021/04/23 by Shahin Amiriparian, Tobias Hübner, Amiriparian, Shahin +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  28. Robust Federated Learning Against Adversarial Attacks for Speech Emotion Recognition
    2022/03/09 by Yi Chang, Sofiane Laridi, Chang, Yi +9 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  29. Adversarial Training in Affective Computing and Sentiment Analysis: Recent Advances and Perspectives
    2018/09/21 by Jing Han, Han, Jing, Zixing Zhang +5 · 1 citation
    Computer Science · #Sentiment Analysis and Opinion Mining #Generative Adversarial Networks and Image Synthesis #Anomaly Detection Techniques and Applications
  30. Augmenting Generative Adversarial Networks for Speech Emotion\n Recognition
    2020/05/18 by Siddique Latif, Latif, Siddique, Muhammad Asim +9 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  31. COLD Fusion: Calibrated and Ordinal Latent Distribution Fusion for Uncertainty-Aware Multimodal Emotion Recognition
    2022/06/12 by Mani Kumar Tellamekala, Shahin Amiriparian, Tellamekala, Mani Kumar +9 · 1 citation
    Computer Science · Psychology · #Computer Vision and Pattern Recognition (cs.CV) #Emotion and Mood Recognition #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Multimedia (cs.MM) #Music and Audio Processing #Speech and Audio Processing
  32. A Comprehensive Survey on Heart Sound Analysis in the Deep Learning Era
    2023/01/23 by Zhao Ren, Ren, Zhao, Yi Chang +9 · 1 citation
    Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Phonocardiography and Auscultation Techniques #Sound (cs.SD) #electronic engineering #information engineering
  33. An Overview & Analysis of Sequence-to-Sequence Emotional Voice Conversion
    2022/03/29 by Zijiang Yang, Xin Jing, Yang, Zijiang +9 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis
  34. ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
    2024/12/16 by Xiangheng He, Junjie Chen, He, Xiangheng +5 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  35. Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT
    2023/03/03 by Mostafa M. Amin, Amin, Mostafa M., Erik Cambria +3 · 2 citations
    Psychology · Computer Science · #Mental Health via Writing #Sentiment Analysis and Opinion Mining #Emotion and Mood Recognition
  36. On Prompt Sensitivity of ChatGPT in Affective Computing
    2024/03/20 by Mostafa M. Amin, Björn W. Schuller, Amin, Mostafa M. +1 · 1 citation
    Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  37. Heart Sound Abnormality Detection From Multi-Institutional Collaboration: Introducing a Federated Learning Framework
    2024/05/03 by Wanyong Qiu, Quan Chen, Chen Quan +10 · 1 citation
    Medicine · Health Professions · Computer Science · #Phonocardiography and Auscultation Techniques #Artificial Intelligence in Healthcare #Machine Learning in Healthcare
  38. The ICSTM+TUM+UP Approach to the 3rd CHIME Challenge: Single-Channel\n LSTM Speech Enhancement with Multi-Channel Correlation Shaping\n Dereverberation and LSTM Language Models
    2015/10/01 by Amr El-Desoky Mousa, Erik Marchi, Mousa, Amr El-Desoky +3 · 1 citation
    Computer Science · Engineering · #Speech and Audio Processing #Speech Recognition and Synthesis #Advanced Adaptive Filtering Techniques
  39. Audio Explanation Synthesis with Generative Foundation Models
    2024/10/10 by Alican Akman, Qiyang Sun, Akman, Alican +3 · 1 citation
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
  40. DFingerNet: Noise-Adaptive Speech Enhancement for Hearing Aids
    2025/01/17 by Iosif Tsangko, Andreas Triantafyllopoulos, Tsangko, Iosif +7 · 1 citation
    Computer Science · Engineering · Neuroscience · #Speech and Audio Processing #Advanced Adaptive Filtering Techniques #Hearing Loss and Rehabilitation
  41. DOTA-ME-CS: Daily Oriented Text Audio-Mandarin English-Code Switching Dataset
    2025/01/21 by Yupei Li, Li, Yupei, Wei, Zifan +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques