Tian Liang
- Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
2024/12/30 by Xingyu Chen, Jiahao Xu, Chen, Xingyu +27 · 2 voices · 121 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.CL #semigroups and automata theory
- Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
2025/01/30 by Yue Wang, Qiuzhi Liu, Wang, Yue +27 · 4 voices · 39 citations
Computer Science · Social Sciences · #Artificial Intelligence in Law #Digital Rights Management and Security #cs.CL
- DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
2025/04/15 by Zhiwei He, Tian Liang, He, Zhiwei +27 · 72 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning and Data Classification #Multimodal Machine Learning Applications #Topic Modeling
- How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
2024/03/18 by Jen-tse Huang, Eric John Li, Huang, Jen-tse +17 · 6 citations
Business, Management and Accounting · Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FinTech, Crowdfunding, Digital Finance #Franchising Strategies and Performance #Open Source Software Innovations
- Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards
2025/05/19 by Xiaoyuan Liu, Tian Liang, Liu, Xiaoyuan +15 · 14 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Topic Modeling
- Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training
2025/05/20 by Wang, Mengru, Xingyu Chen, Chen, Xingyu +25 · 8 citations
Computer Science · #Cognitive Science and Mapping #AI-based Problem Solving and Planning
- FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention
2026/06/08 by Yan Wang, Qifan Zhang, Jiachen Yu +12 · 2 voices · 1 citation
#cs.LG #cs.AI
- A new high-order shock-capturing TENO scheme combined with skew-symmetric-splitting method for compressible gas dynamics and turbulence simulation
2024/05/10 by Tian Liang, Lin Fu · 1 citation
Engineering · Mathematics · #Computational Fluid Dynamics and Aerodynamics #Gas Dynamics and Kinetic Theory #Fluid Dynamics and Turbulent Flows
- DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning
2025/05/29 by Ziyin Zhang, Zhang, Ziyin, Jiahao Xu +23 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Logic, programming, and type systems #Multi-Agent Systems and Negotiation #Software Engineering Research