Jin, Yizhu
- Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
2025/02/06 by Ye, Zhen, Zhu, Xinfa, Chan, Chi-Min +17 · 38 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
- AudioX: A Unified Framework for Anything-to-Audio Generation
2025/03/13 by Zhaoyang Liu, Tian, Zeyue, Jin, Yizhu +10 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- I-MedSAM: Implicit Medical Image Segmentation with Segment Anything
2023/11/28 by Xiaobao Wei, Jiajun Cao, Wei, Xiaobao +9 · 3 citations
Computer Science · Medicine · #AI in cancer detection #COVID-19 diagnosis using AI #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Radiomics and Machine Learning in Medical Imaging
- A Dataset and Benchmark for Copyright Infringement Unlearning from Text-to-Image Diffusion Models
2024/01/04 by Ma, Rui, Zhou, Qiang, Jin, Yizhu +11 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences