vix.ing · top · new · best · stats · spec

Yong Ren

  1. ADD 2023: the Second Audio Deepfake Detection Challenge
    2023/05/23 by Jiangyan Yi, Yi, Jiangyan, Jianhua Tao +33 · 25 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Digital Media Forensic Detection #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  2. Conditional Generative Moment-Matching Networks
    2016/06/14 by Yong Ren, Ren, Yong, Jialian Li +5 · 4 citations
    Computer Science · #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning in Healthcare #Topic Modeling
  3. Step-Audio 2 Technical Report
    2025/07/22 by Boyong Wu, Chao Yan, Wu, Boyong +194 · 33 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  4. AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models
    2025/01/27 by Zheng Lian, Lian, Zheng, Haoyu Chen +21 · 14 citations
    Computer Science · #Sentiment Analysis and Opinion Mining
  5. MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
    2024/01/07 by Zheng Lian, Licai Sun, Lian, Zheng +13 · 7 citations
    Psychology · Computer Science · #Emotion and Mood Recognition #Sentiment Analysis and Opinion Mining #Text and Document Classification Technologies
  6. Fewer-token Neural Speech Codec with Time-invariant Codes
    2023/09/15 by Yong Ren, Ren, Yong, Tao Wang +11 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Video-to-Audio Generation with Hidden Alignment
    2024/07/10 by Manjie Xu, Xu, Manjie, Chenxing Li +12 · 7 citations
    Computer Science · #Music Technology and Sound Studies
  8. Age of Information in Energy Harvesting Aided Massive Multiple Access Networks
    2021/12/23 by Zhengru Fang, Fang, Zhengru, Jingjing Wang +8 · 2 citations
    Computer Science · Engineering · #Age of Information Optimization #FOS: Computer and information sciences #Information Theory (cs.IT) #IoT Networks and Protocols #IoT and Edge/Fog Computing
  9. Environment-Aware AUV Trajectory Design and Resource Management for Multi-Tier Underwater Computing
    2022/10/26 by Xiangwang Hou, Hou, Xiangwang, Jingjing Wang +8 · 2 citations
    Computer Science · Engineering · #Distributed #FOS: Computer and information sciences #IoT and Edge/Fog Computing #Maritime Navigation and Safety #Parallel #Underwater Vehicles and Communication Systems #and Cluster Computing (cs.DC)
  10. Region-Based Optimization in Continual Learning for Audio Deepfake Detection
    2024/12/16 by Yujie Chen, Chen, Yujie, Jiangyan Yi +23 · 4 citations
    Computer Science · #Anomaly Detection Techniques and Applications #Digital Media Forensic Detection #Speech and Audio Processing
  11. STA-V2A: Video-to-Audio Generation with Semantic and Temporal Alignment
    2024/09/13 by Yong Ren, Chenxing Li, Ren, Yong +11 · 2 citations
    Computer Science · #Music and Audio Processing #Music Technology and Sound Studies
  12. ALLM4ADD: Unlocking the Capabilities of Audio Large Language Models for Audio Deepfake Detection
    2025/05/16 by Hao Gu, Jiangyan Yi, Gu, Hao +15 · 3 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Speech Recognition and Synthesis #Music and Audio Processing
  13. SpikeVoice: High-Quality Text-to-Speech Via Efficient Spiking Neural Network
    2024/07/17 by Kexin Wang, Wang, Kexin, Jiahong Zhang +11 · 1 citation
    Computer Science · Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Robotics and Automated Systems #Speech Recognition and Synthesis
  14. Towards Diverse and Efficient Audio Captioning via Diffusion Models
    2024/09/14 by Manjie Xu, Chenxing Li, Xu, Manjie +11 · 1 citation
    Arts and Humanities · Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech and Audio Processing #Subtitles and Audiovisual Media
  15. Evaluating Large Language Models on Financial Report Summarization: An Empirical Study
    2024/11/11 by Yang Xinqi, Yang, Xinqi, Scott Zang +7 · 1 citation
    Computer Science · Decision Sciences · #Advanced Text Analysis Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Stock Market Forecasting Methods #Topic Modeling
  16. Reject Threshold Adaptation for Open-Set Model Attribution of Deepfake Audio
    2024/12/02 by Yan, Xinrui, Jiangyan Yi, Jianhua Tao +13 · 1 citation
    Computer Science · #Speech and Audio Processing #Image and Signal Denoising Methods #Music and Audio Processing