vix.ing · top · new · best · stats · spec

Anikait Singh

  1. Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
    2025/03/03 by Kanishk Gandhi, Ayush Chakravarthy, Gandhi, Kanishk +7 · 22 voices · 127 citations
    #cs.CL #cs.LG
  2. Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
    2025/01/08 by Violet Xiang, Charlie Snell, Xiang, Violet +25 · 17 voices · 17 citations
    #cs.AI #cs.CL
  3. RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
    2023/07/28 by Anthony Brohan, Noah Brown, Brohan, Anthony +105 · 561 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Robotics (cs.RO) #Topic Modeling
  4. Open X-Embodiment: Robotic Learning Datasets and RT-X Models
    2023/10/13 by Embodiment Collaboration, O'Neill, Abby, Rehman, Abdul +349 · 175 citations
    Computer Science · Engineering · #FOS: Computer and information sciences #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO)
  5. Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
    2023/03/09 by Mitsuhiko Nakamoto, Nakamoto, Mitsuhiko, Yuexiang Zhai +13 · 33 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics #Software Engineering Research
  6. Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
    2024/04/22 by Fahim Tajwar, Tajwar, Fahim, Anikait Singh +15 · 26 citations
    Decision Sciences · Economics, Econometrics and Finance · #Efficiency Analysis Using DEA #Healthcare Policy and Management #Auction Theory and Applications
  7. Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
    2025/02/24 by Alon Albalak, Albalak, Alon, Duy Phung +18 · 29 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  8. When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?
    2022/04/12 by Aviral Kumar, Joey Hong, Kumar, Aviral +5 · 7 citations
    Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #Reinforcement Learning in Robotics
  9. A Workflow for Offline Model-Free Robotic Reinforcement Learning
    2021/09/22 by Aviral Kumar, Anikait Singh, Kumar, Aviral +7 · 3 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Machine Learning and Algorithms #Advanced Bandit Algorithms Research
  10. RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
    2025/10/02 by Yuzhong Qu, Anikait Singh, Qu, Yuxiao +11 · 4 citations
    Social Sciences · Computer Science · #Artificial Intelligence in Law #Software Engineering Research #Natural Language Processing Techniques
  11. D5RL: Diverse Datasets for Data-Driven Deep Reinforcement Learning
    2024/08/15 by Rafael Rafailov, Kyle Hatch, Rafailov, Rafael +21 · 2 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  12. MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline
    2025/10/08 by Rushi Qiang, Qiang, Rushi, Yuchen Zhuang +11 · 3 citations
    Computer Science · #Natural Language Processing Techniques
  13. Test-Time Alignment via Hypothesis Reweighting
    2024/12/11 by Yoonho Lee, Lee, Yoonho, Jonathan S. Williams +11 · 1 citation
    Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Machine Learning (cs.LG)