Higuchi, Takuya
- ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models
2024/01/30 by Jee-weon Jung, Jung, Jee-weon, Wangyou Zhang +13 · 9 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Can you Remove the Downstream Model for Speaker Recognition with Self-Supervised Speech Features?
2024/02/01 by Zakaria Aldeneh, Aldeneh, Zakaria, Takuya Higuchi +15 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Multichannel Voice Trigger Detection Based on Transform-average-concatenate
2023/09/27 by Takuya Higuchi, Higuchi, Takuya, Avamarie Brueggeman +5 · 2 citations
Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Sub-cycle temporal evolution of light-induced electron dynamics in\n hexagonal 2D materials
2020/02/13 by Christian Heide, Tobias Boolakee, Heide, Christian +5 · 1 citation
Engineering · Physics and Astronomy · #FOS: Physical sciences #Mesoscale and Nanoscale Physics (cond-mat.mes-hall) #Optics (physics.optics) #Photocathodes and Microchannel Plates #Semiconductor Quantum Structures and Devices #Spectroscopy and Quantum Chemical Studies
- Towards Automatic Assessment of Self-Supervised Speech Models using Rank
2024/09/16 by Aldeneh, Zakaria, Thilak, Vimal, Higuchi, Takuya +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- A Variational Framework for Improving Naturalness in Generative Spoken Language Models
2025/06/17 by Liwei Chen, Chen, Li-Wei, Takuya Higuchi +7 · 2 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems