Changxin Tian
- Every FLOP Counts: Scaling a 300B Mixture-of-Experts LING LLM without Premium GPUs
2025/03/07 by Ling Team, Binwei Zeng, Zeng, Binwei +144 · 10 voices · 5 citations
#cs.LG #cs.AI #cs.CL
- Towards Greater Leverage: Scaling Laws for Efficient Mixture-of-Experts Language Models
2025/07/23 by Changxin Tian, Tian, Changxin, Kunlong Chen +9 · 2 voices · 8 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #I.2.7 #Machine Learning in Healthcare #Speech and dialogue systems #Topic Modeling
- Privacy-preserving Cross-domain Recommendation with Federated Graph Learning
2024/05/13 by Changxin Tian, Yuexiang Xie, Xu Chen +2 · 2 citations
- WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training
2025/07/23 by Changxin Tian, Tian, Changxin, Jiapeng Wang +17 · 5 citations
Computer Science · Engineering · #Analog and Mixed-Signal Circuit Design #Computation and Language (cs.CL) #Distributed systems and fault tolerance #FOS: Computer and information sciences #I.2.7 #Machine Learning (cs.LG) #VLSI and Analog Circuit Testing