Song, Xiaoniu
- Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
2025/02/14 by Guoqing Ma, Haoyang Huang, Ma, Guoqing +235 · 4 voices · 60 citations
Computer Science · #Video Coding and Compression Technologies #cs.CL #cs.CV
- StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
2025/04/22 by Yinmin Zhong, Zhong, Yinmin, Zili Zhang +25 · 18 citations
Engineering · Computer Science · #VLSI and FPGA Design Techniques #Iterative Learning Control Systems #Algorithms and Data Compression
- ProMoE: Fast MoE-based LLM Serving using Proactive Caching
2024/10/29 by Xiaoniu Song, Zihang Zhong, Song, Xiaoniu +4 · 10 citations
Computer Science · #Caching and Content Delivery #Mobile Ad Hoc Networks #Cooperative Communication and Network Coding
- Step-3 is Large yet Affordable: Model-system Co-design for Cost-effective Decoding
2025/07/25 by StepFun, :, Bin Wang +284 · 18 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Step-Video-TI2V Technical Report: A State-of-the-Art Text-Driven Image-to-Video Generation Model
2025/03/14 by Haoyang Huang, Guoqing Ma, Huang, Haoyang +100 · 6 citations
Computer Science · #Advanced Vision and Imaging #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications