Anikait Singh
- Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
2025/03/03 by Kanishk Gandhi, Ayush Chakravarthy, Gandhi, Kanishk +7 · 22 voices · 127 citations
#cs.CL #cs.LG
- Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
2025/01/08 by Violet Xiang, Charlie Snell, Xiang, Violet +25 · 17 voices · 17 citations
#cs.AI #cs.CL
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
2023/07/28 by Anthony Brohan, Noah Brown, Brohan, Anthony +105 · 561 citations
Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Robotics (cs.RO) #Topic Modeling
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models
2023/10/13 by Embodiment Collaboration, O'Neill, Abby, Rehman, Abdul +349 · 175 citations
Computer Science · Engineering · #FOS: Computer and information sciences #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO)
- Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
2023/03/09 by Mitsuhiko Nakamoto, Nakamoto, Mitsuhiko, Yuexiang Zhai +13 · 33 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics #Software Engineering Research
- Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
2024/04/22 by Fahim Tajwar, Tajwar, Fahim, Anikait Singh +15 · 26 citations
Decision Sciences · Economics, Econometrics and Finance · #Efficiency Analysis Using DEA #Healthcare Policy and Management #Auction Theory and Applications
- Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
2025/02/24 by Alon Albalak, Albalak, Alon, Duy Phung +18 · 29 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?
2022/04/12 by Aviral Kumar, Joey Hong, Kumar, Aviral +5 · 7 citations
Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #Reinforcement Learning in Robotics
- A Workflow for Offline Model-Free Robotic Reinforcement Learning
2021/09/22 by Aviral Kumar, Anikait Singh, Kumar, Aviral +7 · 3 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Machine Learning and Algorithms #Advanced Bandit Algorithms Research
- RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
2025/10/02 by Yuzhong Qu, Anikait Singh, Qu, Yuxiao +11 · 4 citations
Social Sciences · Computer Science · #Artificial Intelligence in Law #Software Engineering Research #Natural Language Processing Techniques
- D5RL: Diverse Datasets for Data-Driven Deep Reinforcement Learning
2024/08/15 by Rafael Rafailov, Kyle Hatch, Rafailov, Rafael +21 · 2 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
- MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline
2025/10/08 by Rushi Qiang, Qiang, Rushi, Yuchen Zhuang +11 · 3 citations
Computer Science · #Natural Language Processing Techniques
- Test-Time Alignment via Hypothesis Reweighting
2024/12/11 by Yoonho Lee, Lee, Yoonho, Jonathan S. Williams +11 · 1 citation
Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Machine Learning (cs.LG)