Bowen Wu
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1866 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models
2025/08/08 by GLM-4. 5 Team, Team, 5 Team +357 · 21 voices · 133 citations
Computer Science · #Cognitive Computing and Networks #Distributed and Parallel Computing Systems #Rough Sets and Fuzzy Logic #cs.CL
- F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization
2025/04/03 by Xiaohui Sun, Ruitong Xiao, Sun, Xiaohui +9 · 12 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- Terabyte-Scale Analytics in the Blink of an Eye
2025/06/10 by Bowen Wu, Wei Cui, Wu, Bowen +7 · 2 voices · 2 citations
Computer Science · #Databases (cs.DB) #Distributed #FOS: Computer and information sciences #Parallel #Performance (cs.PF) #and Cluster Computing (cs.DC) #cs.DB #cs.DC #cs.PF
- RAIDEN-R1: Improving Role-awareness of LLMs via GRPO with Verifiable Reward
2025/05/15 by Z. Yan Wang, K. T. Sun, Wang, Zongsheng +9 · 5 citations
Computer Science · Business, Management and Accounting · #Service-Oriented Architecture and Web Services #Business Process Modeling and Analysis #Software System Performance and Reliability
- Guiding Variational Response Generator to Exploit Persona
2019/11/06 by Bowen Wu, Wu, Bowen, Mengyuan Li +13 · 1 citation
Computer Science · #AI in Service Interactions #Computation and Language (cs.CL) #FOS: Computer and information sciences #Persona Design and Applications #Topic Modeling
- CodegenBench: Can LLMs Write Efficient Code Across Architectures?
2026/06/01 by Jie Li, Wenzhao Wu, Junqi Hu +5 · 2 voices
Computer Science · #cs.SE #cs.AI
- From Sign Language Generation to Humanoid Execution: Vision-Language Guided Retargeting with Collision Mitigation
2026/07/20 by Nabeela Khan, Bowen Wu, Runwu Shi +5
#cs.RO #cs.CV #cs.HC