Shrikanth Narayanan
- A Review of Speaker Diarization: Recent Advances with Deep Learning
2021/01/24 by Tae Jin Park, Park, Tae Jin, Naoyuki Kanda +9 · 17 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- FedMultimodal: A Benchmark For Multimodal Federated Learning
2023/06/15 by Tiantian Feng, Feng, Tiantian, Digbalay Bose +15 · 9 citations
Computer Science · #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel #Privacy-Preserving Technologies in Data #and Cluster Computing (cs.DC)
- Paralinguistics in speech and language—State-of-the-art and the challenge
2012/03/08 by Björn W. Schuller, Björn Schuller, Stefan Steidl +5 · 5 citations
Psychology · #Emotion and Mood Recognition #Multisensory perception and integration #Phonetics and Phonology Research
- End-to-End Neural Systems for Automatic Children Speech Recognition: An\n Empirical Study
2021/02/19 by Prashanth Gurunath Shivakumar, Shrikanth Narayanan, Shivakumar, Prashanth Gurunath +1 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- FedAudio: A Federated Learning Benchmark for Audio Tasks
2022/10/27 by Tuo Zhang, Zhang, Tuo, Tiantian Feng +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Distributed #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Parallel #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #and Cluster Computing (cs.DC) #electronic engineering #information engineering
- Creating a Lens of Chinese Culture: A Multimodal Dataset for Chinese Pun Rebus Art Understanding
2024/06/14 by Tuo Zhang, Tiantian Feng, Zhang, Tuo +17 · 4 citations
Arts and Humanities · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cultural Heritage Management and Preservation #FOS: Computer and information sciences
- Characterizing Types of Convolution in Deep Convolutional Recurrent Neural Networks for Robust Speech Emotion Recognition
2017/06/07 by Che-Wei Huang, Huang, Che-Wei, Shrikanth Narayanan +1 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing
- MM-AU:Towards Multimodal Understanding of Advertisement Videos
2023/08/27 by Digbalay Bose, Rajat Hebbar, Bose, Digbalay +9 · 2 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies #Topic Modeling
- LSM-2: Learning from Incomplete Wearable Sensor Data
2025/06/05 by Maxwell A. Xu, Xu, Maxwell A., Girish Narayanswamy +47 · 6 citations
Computer Science · Engineering · #Context-Aware Activity Recognition Systems #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Non-Invasive Vital Sign Monitoring
- The INTERSPEECH 2020 Far-Field Speaker Verification Challenge
2020/05/16 by Xiaoyi Qin, Qin, Xiaoyi, Ming Li +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Understanding of Emotion Perception from Art
2021/10/13 by Digbalay Bose, Bose, Digbalay, Krishna Somandepalli +9 · 1 citation
Computer Science · Neuroscience · #Multimodal Machine Learning Applications #Aesthetic Perception and Analysis #Generative Adversarial Networks and Image Synthesis
- Audio-Visual Activity Guided Cross-Modal Identity Association for Active Speaker Detection
2022/12/01 by Rahul Sharma, Shrikanth Narayanan, Sharma, Rahul +1 · 1 citation
Arts and Humanities · Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM) #Music and Audio Processing #Speech and Audio Processing #Subtitles and Audiovisual Media
- TrustSER: On the Trustworthiness of Fine-tuning Pre-trained Speech Embeddings For Speech Emotion Recognition
2023/05/18 by Tiantian Feng, Feng, Tiantian, Rajat Hebbar +3 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Emotion-Aligned Contrastive Learning Between Images and Music
2023/08/24 by Shanti Stewart, Kleanthis Avramidis, Stewart, Shanti +5 · 1 citation
Arts and Humanities · Computer Science · #Audio and Speech Processing (eess.AS) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Speech2rtMRI: Speech-Guided Diffusion Model for Real-time MRI Video of the Vocal Tract during Speech
2024/09/23 by Hong Nguyen, Sean Foley, Nguyen, Hong +9 · 2 citations
Computer Science · #Speech Recognition and Synthesis
- GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
2023/06/03 by Tuo Zhang, Zhang, Tuo, Tiantian Feng +11 · 1 citation
Computer Science · #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Parallel #Privacy-Preserving Technologies in Data #and Cluster Computing (cs.DC)
- CHATTER: A Character Attribution Dataset for Narrative Understanding
2024/11/07 by Sabyasachee Baruah, Shrikanth Narayanan, Baruah, Sabyasachee +1 · 1 citation
Computer Science · Social Sciences · #Topic Modeling #Natural Language Processing Techniques #Computational and Text Analysis Methods
- Leveraging Label Correlations in a Multi-label Setting: A Case Study in Emotion
2022/10/28 by Georgios Chochlakis, Chochlakis, Georgios, Gireesh Mahajan +9 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies #Topic Modeling
- Using Emotion Embeddings to Transfer Knowledge Between Emotions, Languages, and Annotation Formats
2022/10/31 by Georgios Chochlakis, Chochlakis, Georgios, Gireesh Mahajan +9 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Sentiment Analysis and Opinion Mining #Topic Modeling
- Knowledge-guided EEG Representation Learning
2024/02/15 by Aditya Kommineni, Kleanthis Avramidis, Kommineni, Aditya +5 · 1 citation
Computer Science · Neuroscience · #Artificial Intelligence (cs.AI) #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural Networks and Applications #Signal Processing (eess.SP) #electronic engineering #information engineering
- Speech Entrainment in Multi-Party Conversations with a Digital Agent
2026/07/24 by Nicholas Mehlman, Kaitlin Zareno, Kleanthis Avramidis +2
#eess.AS
- AMECxSV: Adaptive Metadata-Driven Embedding-Fusion Calibration for X-Lingual Speaker Verification
2026/07/17 by Xin Wei, Shi He, Yihe Yuan +3
#eess.AS