vix.ing · top · new · best · stats · spec

Shrikanth Narayanan

  1. A Review of Speaker Diarization: Recent Advances with Deep Learning
    2021/01/24 by Tae Jin Park, Park, Tae Jin, Naoyuki Kanda +9 · 17 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  2. FedMultimodal: A Benchmark For Multimodal Federated Learning
    2023/06/15 by Tiantian Feng, Feng, Tiantian, Digbalay Bose +15 · 9 citations
    Computer Science · #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel #Privacy-Preserving Technologies in Data #and Cluster Computing (cs.DC)
  3. Paralinguistics in speech and language—State-of-the-art and the challenge
    2012/03/08 by Björn W. Schuller, Björn Schuller, Stefan Steidl +5 · 5 citations
    Psychology · #Emotion and Mood Recognition #Multisensory perception and integration #Phonetics and Phonology Research
  4. End-to-End Neural Systems for Automatic Children Speech Recognition: An\n Empirical Study
    2021/02/19 by Prashanth Gurunath Shivakumar, Shrikanth Narayanan, Shivakumar, Prashanth Gurunath +1 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. FedAudio: A Federated Learning Benchmark for Audio Tasks
    2022/10/27 by Tuo Zhang, Zhang, Tuo, Tiantian Feng +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Distributed #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Parallel #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #and Cluster Computing (cs.DC) #electronic engineering #information engineering
  6. Creating a Lens of Chinese Culture: A Multimodal Dataset for Chinese Pun Rebus Art Understanding
    2024/06/14 by Tuo Zhang, Tiantian Feng, Zhang, Tuo +17 · 4 citations
    Arts and Humanities · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cultural Heritage Management and Preservation #FOS: Computer and information sciences
  7. Characterizing Types of Convolution in Deep Convolutional Recurrent Neural Networks for Robust Speech Emotion Recognition
    2017/06/07 by Che-Wei Huang, Huang, Che-Wei, Shrikanth Narayanan +1 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing
  8. MM-AU:Towards Multimodal Understanding of Advertisement Videos
    2023/08/27 by Digbalay Bose, Rajat Hebbar, Bose, Digbalay +9 · 2 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies #Topic Modeling
  9. LSM-2: Learning from Incomplete Wearable Sensor Data
    2025/06/05 by Maxwell A. Xu, Xu, Maxwell A., Girish Narayanswamy +47 · 6 citations
    Computer Science · Engineering · #Context-Aware Activity Recognition Systems #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Non-Invasive Vital Sign Monitoring
  10. The INTERSPEECH 2020 Far-Field Speaker Verification Challenge
    2020/05/16 by Xiaoyi Qin, Qin, Xiaoyi, Ming Li +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Understanding of Emotion Perception from Art
    2021/10/13 by Digbalay Bose, Bose, Digbalay, Krishna Somandepalli +9 · 1 citation
    Computer Science · Neuroscience · #Multimodal Machine Learning Applications #Aesthetic Perception and Analysis #Generative Adversarial Networks and Image Synthesis
  12. Audio-Visual Activity Guided Cross-Modal Identity Association for Active Speaker Detection
    2022/12/01 by Rahul Sharma, Shrikanth Narayanan, Sharma, Rahul +1 · 1 citation
    Arts and Humanities · Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM) #Music and Audio Processing #Speech and Audio Processing #Subtitles and Audiovisual Media
  13. TrustSER: On the Trustworthiness of Fine-tuning Pre-trained Speech Embeddings For Speech Emotion Recognition
    2023/05/18 by Tiantian Feng, Feng, Tiantian, Rajat Hebbar +3 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  14. Emotion-Aligned Contrastive Learning Between Images and Music
    2023/08/24 by Shanti Stewart, Kleanthis Avramidis, Stewart, Shanti +5 · 1 citation
    Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  15. Speech2rtMRI: Speech-Guided Diffusion Model for Real-time MRI Video of the Vocal Tract during Speech
    2024/09/23 by Hong Nguyen, Sean Foley, Nguyen, Hong +9 · 2 citations
    Computer Science · #Speech Recognition and Synthesis
  16. GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
    2023/06/03 by Tuo Zhang, Zhang, Tuo, Tiantian Feng +11 · 1 citation
    Computer Science · #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Parallel #Privacy-Preserving Technologies in Data #and Cluster Computing (cs.DC)
  17. CHATTER: A Character Attribution Dataset for Narrative Understanding
    2024/11/07 by Sabyasachee Baruah, Shrikanth Narayanan, Baruah, Sabyasachee +1 · 1 citation
    Computer Science · Social Sciences · #Topic Modeling #Natural Language Processing Techniques #Computational and Text Analysis Methods
  18. Leveraging Label Correlations in a Multi-label Setting: A Case Study in Emotion
    2022/10/28 by Georgios Chochlakis, Chochlakis, Georgios, Gireesh Mahajan +9 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies #Topic Modeling
  19. Using Emotion Embeddings to Transfer Knowledge Between Emotions, Languages, and Annotation Formats
    2022/10/31 by Georgios Chochlakis, Chochlakis, Georgios, Gireesh Mahajan +9 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Sentiment Analysis and Opinion Mining #Topic Modeling
  20. Knowledge-guided EEG Representation Learning
    2024/02/15 by Aditya Kommineni, Kleanthis Avramidis, Kommineni, Aditya +5 · 1 citation
    Computer Science · Neuroscience · #Artificial Intelligence (cs.AI) #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural Networks and Applications #Signal Processing (eess.SP) #electronic engineering #information engineering
  21. Speech Entrainment in Multi-Party Conversations with a Digital Agent
    2026/07/24 by Nicholas Mehlman, Kaitlin Zareno, Kleanthis Avramidis +2
    #eess.AS
  22. AMECxSV: Adaptive Metadata-Driven Embedding-Fusion Calibration for X-Lingual Speaker Verification
    2026/07/17 by Xin Wei, Shi He, Yihe Yuan +3
    #eess.AS