vix.ing · top · new · best · stats · spec

Sharma, Archit

  1. Direct Preference Optimization: Your Language Model is Secretly a Reward Model
    2023/05/29 by Rafael Rafailov, Archit Sharma, Rafailov, Rafael +9 · 9 voices · 1630 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems
  2. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1389 citations
    #cs.CL #cs.AI
  3. DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
    2024/03/19 by Alexander Khazatsky, Karl Pertsch, Khazatsky, Alexander +196 · 212 citations
    Computer Science · Engineering · #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotic Path Planning Algorithms #Robotics (cs.RO)
  4. Open X-Embodiment: Robotic Learning Datasets and RT-X Models
    2023/10/13 by Embodiment Collaboration, O'Neill, Abby, Rehman, Abdul +349 · 180 citations
    Computer Science · Engineering · #FOS: Computer and information sciences #Modular Robots and Swarm Intelligence #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotics (cs.RO)
  5. Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
    2023/05/24 by Katherine Tian, Eric Mitchell, Tian, Katherine +13 · 123 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
  6. Stream of Search (SoS): Learning to Search in Language
    2024/04/01 by Kanishk Gandhi, Denise Lee, Gandhi, Kanishk +11 · 3 voices · 20 citations
    Computer Science · #Speech and dialogue systems #cs.AI #cs.CL #cs.LG
  7. Dynamics-Aware Unsupervised Discovery of Skills
    2019/07/02 by Archit Sharma, Sharma, Archit, Shixiang Gu +7 · 22 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Robotics (cs.RO)
  8. Yell At Your Robot: Improving On-the-Fly from Language Corrections
    2024/03/19 by Lucy Xiaoyang Shi, Shi, Lucy Xiaoyang, Zheyuan Hu +13 · 31 citations
    Computer Science · Engineering · #AI in Service Interactions #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO) #Robotics and Automated Systems #Speech and dialogue systems
  9. SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning
    2024/01/29 by Luo, Jianlan, Hu, Zheyuan, Xu, Charles +7 · 28 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Robotics (cs.RO)
  10. Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
    2024/04/22 by Fahim Tajwar, Tajwar, Fahim, Anikait Singh +15 · 26 citations
    Decision Sciences · Economics, Econometrics and Finance · #Efficiency Analysis Using DEA #Healthcare Policy and Management #Auction Theory and Applications
  11. Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
    2024/12/09 by Max Sobol Mark, Mark, Max Sobol, Tian Gao +11 · 1 voice · 20 citations
    #cs.LG #cs.AI
  12. Waypoint-Based Imitation Learning for Robotic Manipulation
    2023/07/26 by Lucy Xiaoyang Shi, Archit Sharma, Shi, Lucy Xiaoyang +5 · 10 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Motion and Animation #Machine Learning (cs.LG) #Robot Manipulation and Learning #Robotic Path Planning Algorithms #Robotics (cs.RO)
  13. An Emulator for Fine-Tuning Large Language Models using Small Language Models
    2023/10/19 by Eric Mitchell, Mitchell, Eric, Rafael Rafailov +7 · 10 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Text Readability and Simplification
  14. Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning
    2023/03/02 by Sharma, Archit, Ahmed, Ahmed M., Ahmad, Rehaan +1 · 7 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  15. Robot Fine-Tuning Made Easy: Pre-Training Rewards and Policies for Autonomous Real-World Reinforcement Learning
    2023/10/23 by Yang, Jingyun, Mark, Max Sobol, Vu, Brandon +3 · 8 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  16. Emergent Real-World Robotic Skills via Unsupervised Off-Policy Reinforcement Learning
    2020/04/27 by Archit Sharma, Michael J. Ahn, Sharma, Archit +9 · 4 citations
    Computer Science · Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotic Locomotion and Control #Robotics (cs.RO)
  17. Autonomous Reinforcement Learning: Formalism and Benchmarking
    2021/12/17 by Sharma, Archit, Xu, Kelvin, Sardana, Nikhil +4 · 4 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  18. When to Ask for Help: Proactive Interventions in Autonomous Reinforcement Learning
    2022/10/19 by Xie, Annie, Tajwar, Fahim, Sharma, Archit +1 · 4 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  19. A Critical Evaluation of AI Feedback for Aligning Large Language Models
    2024/02/19 by Sharma, Archit, Keh, Sedrick, Mitchell, Eric +3 · 5 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  20. Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
    2024/10/30 by Hsu, Sheryl, Khattab, Omar, Finn, Chelsea +1 · 6 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  21. FSPO: Few-Shot Preference Optimization of Synthetic Preference Data in LLMs Elicits Effective Personalization to Real Users
    2025/02/26 by Singh, Anikait, Hsu, Sheryl, Hsu, Kyle +5 · 7 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  22. RLVF: Learning from Verbal Feedback without Overgeneralization
    2024/02/16 by Moritz Stephan, Alexander Khazatsky, Stephan, Moritz +11 · 2 citations
    Computer Science · #Fuzzy Logic and Control Systems #Neural Networks and Applications #Intelligent Tutoring Systems and Adaptive Learning
  23. Variational Empowerment as Representation Learning for Goal-Based\n Reinforcement Learning
    2021/06/02 by Jongwook Choi, Choi, Jongwook, Archit Sharma +7 · 1 citation
    Computer Science · Neuroscience · Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mental Health Research Topics #Neural and Behavioral Psychology Studies #Neural dynamics and brain function #Reinforcement Learning in Robotics
  24. Towards Data-Centric RLHF: Simple Metrics for Preference Dataset Comparison
    2024/09/15 by Judy Hanwen Shen, Shen, Judy Hanwen, Archit Sharma +3 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. A State-Distribution Matching Approach to Non-Episodic Reinforcement Learning
    2022/05/11 by Archit Sharma, Rehaan Ahmad, Sharma, Archit +3 · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO) #Smart Grid Energy Management
  26. Test-Time Alignment via Hypothesis Reweighting
    2024/12/11 by Yoonho Lee, Jonathan S. Williams, Lee, Yoonho +11 · 1 citation
    Computer Science · #Educational Technology and Assessment #FOS: Computer and information sciences #Machine Learning (cs.LG)