Jiachen Lian
- SSDM: Scalable Speech Dysfluency Modeling
2024/08/29 by Jiachen Lian, Xuanru Zhou, Lian, Jiachen +15 · 10 citations
Computer Science · Psychology · Medicine · #Speech Recognition and Synthesis #Phonetics and Phonology Research #Voice and Speech Disorders
- Robust Disentangled Variational Speech Representation Learning for Zero-shot Voice Conversion
2022/03/30 by Jiachen Lian, Chunlei Zhang, Lian, Jiachen +3 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- AV-data2vec: Self-supervised Learning of Audio-Visual Speech Representations with Contextualized Target Representations
2023/02/10 by Jiachen Lian, Lian, Jiachen, Alexei Baevski +5 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Stutter-Solver: End-to-end Multi-lingual Dysfluency Detection
2024/09/15 by Xuanru Zhou, Zhou, Xuanru, Cheol Jun Cho +21 · 6 citations
Psychology · #Stuttering Research and Treatment #Phonetics and Phonology Research
- Towards Hierarchical Spoken Language Dysfluency Modeling
2024/01/18 by Jiachen Lian, Gopala Anumanchipalli, Lian, Jiachen +1 · 7 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Time and Tokens: Benchmarking End-to-End Speech Dysfluency Detection
2024/09/20 by Xuanru Zhou, Zhou, Xuanru, Jiachen Lian +23 · 4 citations
Computer Science · Medicine · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Stuttering Research and Treatment #Voice and Speech Disorders #electronic engineering #information engineering
- Towards Improved Zero-shot Voice Conversion with Conditional DSVAE
2022/05/11 by Jiachen Lian, Chunlei Zhang, Lian, Jiachen +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Seamless Dysfluent Speech Text Alignment for Disordered Speech Analysis
2025/06/05 by Zongli Ye, Ye, Zongli, Jiachen Lian +29 · 6 citations
Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
2024/03/01 by Weiwei Lin, Chenhang He, Lin, Weiwei +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Dysfluent WFST: A Framework for Zero-Shot Speech Dysfluency Transcription and Detection
2025/05/22 by Guo, Chenxu, Jiachen Lian, Xuanru Zhou +27 · 5 citations
Computer Science · Medicine · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Stuttering Research and Treatment #Voice and Speech Disorders #electronic engineering #information engineering
- SSDM 2.0: Time-Accurate Speech Rich Transcription with Non-Fluencies
2024/11/29 by Jiachen Lian, Lian, Jiachen, Xuanru Zhou +15 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques