Dakota Mahan
- Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
2025/01/08 by Violet Xiang, Charlie Snell, Xiang, Violet +25 · 17 voices · 25 citations
#cs.AI #cs.CL
- Generative Reward Models
2024/10/02 by Dakota Mahan, Mahan, Dakota, Duy Phung +14 · 37 citations
Economics, Econometrics and Finance · #Diverse Scientific and Economic Studies #Diverse Specialized Academic Research #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
2025/02/24 by Alon Albalak, Duy Phung, Albalak, Alon +18 · 35 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Surveying the Effects of Quality, Diversity, and Complexity in Synthetic Data From Large Language Models
2024/12/04 by Alex Havrilla, Andrew M. Dai, Havrilla, Alex +36 · 12 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Stable LM 2 1.6B Technical Report
2024/02/27 by Marco Bellagente, Jonathan Tow, Bellagente, Marco +35 · 6 citations
Engineering · #Particle accelerators and beam dynamics #Induction Heating and Inverter Technology #Microwave Engineering and Waveguides
- Hermes 4 Technical Report
2025/08/25 by Ryan Teknium, Roger Jin, Teknium, Ryan +15 · 3 voices · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #cs.AI