Sharma, Archit
- Direct Preference Optimization: Your Language Model is Secretly a Reward Model
2023/05/29 by Rafael Rafailov, Archit Sharma, Rafailov, Rafael +9 · 9 voices · 1630 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1389 citations
#cs.CL #cs.AI
- DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
2024/03/19 by Alexander Khazatsky, Karl Pertsch, Khazatsky, Alexander +196 · 212 citations
Computer Science · Engineering · #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotic Path Planning Algorithms #Robotics (cs.RO)
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models
2023/10/13 by Embodiment Collaboration, O'Neill, Abby, Rehman, Abdul +349 · 180 citations
Computer Science · Engineering · #FOS: Computer and information sciences #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO)
- Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
2023/05/24 by Katherine Tian, Eric Mitchell, Tian, Katherine +13 · 123 citations
Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
- Stream of Search (SoS): Learning to Search in Language
2024/04/01 by Kanishk Gandhi, Denise Lee, Gandhi, Kanishk +11 · 3 voices · 20 citations
Computer Science · #Speech and dialogue systems #cs.AI #cs.CL #cs.LG
- Dynamics-Aware Unsupervised Discovery of Skills
2019/07/02 by Archit Sharma, Sharma, Archit, Shixiang Gu +7 · 22 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Robotics (cs.RO)
- Yell At Your Robot: Improving On-the-Fly from Language Corrections
2024/03/19 by Lucy Xiaoyang Shi, Shi, Lucy Xiaoyang, Zheyuan Hu +13 · 31 citations
Computer Science · Engineering · #AI in Service Interactions #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO) #Robotics and Automated Systems #Speech and dialogue systems
- SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning
2024/01/29 by Luo, Jianlan, Hu, Zheyuan, Xu, Charles +7 · 28 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Robotics (cs.RO)
- Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
2024/04/22 by Fahim Tajwar, Tajwar, Fahim, Anikait Singh +15 · 26 citations
Decision Sciences · Economics, Econometrics and Finance · #Efficiency Analysis Using DEA #Healthcare Policy and Management #Auction Theory and Applications
- Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
2024/12/09 by Max Sobol Mark, Mark, Max Sobol, Tian Gao +11 · 1 voice · 20 citations
#cs.LG #cs.AI
- Waypoint-Based Imitation Learning for Robotic Manipulation
2023/07/26 by Lucy Xiaoyang Shi, Archit Sharma, Shi, Lucy Xiaoyang +5 · 10 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Motion and Animation #Machine Learning (cs.LG) #Robot Manipulation and Learning #Robotic Path Planning Algorithms #Robotics (cs.RO)
- An Emulator for Fine-Tuning Large Language Models using Small Language Models
2023/10/19 by Eric Mitchell, Mitchell, Eric, Rafael Rafailov +7 · 10 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Text Readability and Simplification
- Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning
2023/03/02 by Sharma, Archit, Ahmed, Ahmed M., Ahmad, Rehaan +1 · 7 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
- Robot Fine-Tuning Made Easy: Pre-Training Rewards and Policies for Autonomous Real-World Reinforcement Learning
2023/10/23 by Yang, Jingyun, Mark, Max Sobol, Vu, Brandon +3 · 8 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
- Emergent Real-World Robotic Skills via Unsupervised Off-Policy Reinforcement Learning
2020/04/27 by Archit Sharma, Michael J. Ahn, Sharma, Archit +9 · 4 citations
Computer Science · Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotic Locomotion and Control #Robotics (cs.RO)
- Autonomous Reinforcement Learning: Formalism and Benchmarking
2021/12/17 by Sharma, Archit, Xu, Kelvin, Sardana, Nikhil +4 · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
- When to Ask for Help: Proactive Interventions in Autonomous Reinforcement Learning
2022/10/19 by Xie, Annie, Tajwar, Fahim, Sharma, Archit +1 · 4 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- A Critical Evaluation of AI Feedback for Aligning Large Language Models
2024/02/19 by Sharma, Archit, Keh, Sedrick, Mitchell, Eric +3 · 5 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
2024/10/30 by Hsu, Sheryl, Khattab, Omar, Finn, Chelsea +1 · 6 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- FSPO: Few-Shot Preference Optimization of Synthetic Preference Data in LLMs Elicits Effective Personalization to Real Users
2025/02/26 by Singh, Anikait, Hsu, Sheryl, Hsu, Kyle +5 · 7 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- RLVF: Learning from Verbal Feedback without Overgeneralization
2024/02/16 by Moritz Stephan, Alexander Khazatsky, Stephan, Moritz +11 · 2 citations
Computer Science · #Fuzzy Logic and Control Systems #Neural Networks and Applications #Intelligent Tutoring Systems and Adaptive Learning
- Variational Empowerment as Representation Learning for Goal-Based\n Reinforcement Learning
2021/06/02 by Jongwook Choi, Choi, Jongwook, Archit Sharma +7 · 1 citation
Computer Science · Neuroscience · Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mental Health Research Topics #Neural and Behavioral Psychology Studies #Neural dynamics and brain function #Reinforcement Learning in Robotics
- Towards Data-Centric RLHF: Simple Metrics for Preference Dataset Comparison
2024/09/15 by Judy Hanwen Shen, Shen, Judy Hanwen, Archit Sharma +3 · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG)
- A State-Distribution Matching Approach to Non-Episodic Reinforcement Learning
2022/05/11 by Archit Sharma, Rehaan Ahmad, Sharma, Archit +3 · 1 citation
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO) #Smart Grid Energy Management
- Test-Time Alignment via Hypothesis Reweighting
2024/12/11 by Yoonho Lee, Jonathan S. Williams, Lee, Yoonho +11 · 1 citation
Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Machine Learning (cs.LG)