Nobukatsu Hojo
- CycleGAN-VC2: Improved CycleGAN-based Non-parallel Voice Conversion
2019/04/09 by Takuhiro Kaneko, Kaneko, Takuhiro, Hirokazu Kameoka +5 · 10 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- StarGAN-VC2: Rethinking Conditional Methods for StarGAN-Based Voice Conversion
2019/07/29 by Takuhiro Kaneko, Kaneko, Takuhiro, Hirokazu Kameoka +5 · 12 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- ACVAE-VC: Non-parallel many-to-many voice conversion with auxiliary classifier variational autoencoder
2018/08/13 by Hirokazu Kameoka, Kameoka, Hirokazu, Takuhiro Kaneko +5 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind
2025/01/15 by Kazutoshi Shinoda, Shinoda, Kazutoshi, Nobukatsu Hojo +13 · 5 citations
Computer Science · Decision Sciences · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Law #Cognitive Science and Mapping #Complex Systems and Decision Making #Computation and Language (cs.CL) #FOS: Computer and information sciences
- End-to-End Joint Target and Non-Target Speakers ASR
2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- VoiceGrad: Non-Parallel Any-to-Many Voice Conversion with Annealed Langevin Dynamics
2020/10/06 by Hirokazu Kameoka, Kameoka, Hirokazu, Takuhiro Kaneko +7 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Generative adversarial network-based approach to signal reconstruction from magnitude spectrograms
2018/04/06 by K. Oyamada, Oyamada, Keisuke, Hirokazu Kameoka +9 · 1 citation
Computer Science · #Blind Source Separation Techniques #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Signal Processing (eess.SP) #Speech and Audio Processing #electronic engineering #information engineering