vix.ing · top · new · best · stats · spec

Herman Kamper

  1. Participatory Research for Low-resourced Machine Translation: A Case\n Study in African Languages
    2020/10/05 by Wilhelmina Nekoto, Vukosi Marivate, Nekoto, Wilhelmina +92 · 18 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling
  2. Voice Conversion With Just Nearest Neighbors
    2023/05/30 by Matthew Baas, Benjamin van Niekerk, Baas, Matthew +3 · 17 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Masakhane -- Machine Translation For Africa
    2020/03/13 by Iroro Orife, Orife, Iroro, Julia Kreutzer +47 · 10 citations
    Computer Science · Social Sciences · #Natural Language Processing Techniques #Topic Modeling #Wikis in Education and Collaboration
  4. An embedded segmental K-means model for unsupervised segmentation and\n clustering of speech
    2017/03/23 by Herman Kamper, Kamper, Herman, Karen Livescu +3 · 4 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
  5. Pre-training on high-resource speech recognition improves low-resource\n speech-to-text translation
    2018/09/05 by Sameer Bansal, Bansal, Sameer, Herman Kamper +7 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  6. Improved acoustic word embeddings for zero-resource languages using\n multilingual transfer
    2020/06/02 by Herman Kamper, Yevgen Matusevych, Kamper, Herman +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Fast ASR-free and almost zero-resource keyword spotting using DTW and\n CNNs for humanitarian monitoring
    2018/06/25 by Raghav Menon, Menon, Raghav, Herman Kamper +5 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems
  8. Deep convolutional acoustic word embeddings using word-pair side\n information
    2015/10/05 by Herman Kamper, Weiran Wang, Kamper, Herman +3 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling
  9. Query-by-Example Search with Discriminative Neural Acoustic Word\n Embeddings
    2017/06/12 by Shane Settle, Settle, Shane, Keith Levin +5 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
  10. Learning Dynamics of Linear Denoising Autoencoders
    2018/06/14 by Arnu Pretorius, Pretorius, Arnu, Steve Kroon +3 · 1 citation
    Computer Science · Physics and Astronomy · #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks
  11. A Correspondence Variational Autoencoder for Unsupervised Acoustic Word Embeddings
    2020/12/03 by Puyuan Peng, Peng, Puyuan, Herman Kamper +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  12. A comparison of self-supervised speech representations as input features\n for unsupervised acoustic word embeddings
    2020/12/14 by Lisa van Staden, van Staden, Lisa, Herman Kamper +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  13. TransFusion: Transcribing Speech with Multinomial Diffusion
    2022/10/14 by Matthew Baas, Baas, Matthew, Kevin Eloff +3 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  14. Multilingual transfer of acoustic word embeddings improves when training\n on languages related to the target zero-resource language
    2021/06/24 by Jacobs Christiaan, Jacobs, Christiaan, Herman Kamper +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  15. Spoken-Term Discovery using Discrete Speech Units
    2024/08/26 by Benjamin van Niekerk, van Niekerk, Benjamin, Julian Zaïdi +5 · 3 citations
    Computer Science · #Speech and dialogue systems #Natural Language Processing Techniques
  16. Critical initialisation for deep signal propagation in noisy rectifier\n neural networks
    2018/11/01 by Arnu Pretorius, Pretorius, Arnu, Elan van Biljon +5 · 2 citations
    Computer Science · #Machine Learning and ELM #Stochastic Gradient Optimization Techniques #Neural Networks and Applications
  17. MARS6: A Small and Robust Hierarchical-Codec Text-to-Speech Model
    2025/01/10 by Matthew Baas, Pieter Scholtz, Baas, Matthew +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering