Sikchi, Harshit
- gpt-oss-120b & gpt-oss-20b Model Card
2025/08/08 by OpenAI, :, Agarwal, Sandhini +124 · 251 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
2024/06/05 by Rafael Rafailov, Yaswanth Chittepu, Rafailov, Rafael +13 · 16 citations
Engineering · Decision Sciences · Computer Science · #Advanced Control Systems Optimization #Simulation Techniques and Applications #Statistical and Computational Modeling
- Learning Off-Policy with Online Planning
2020/08/23 by Sikchi, Harshit, Zhou, Wenxuan, Held, David · 6 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
- Contrastive Preference Learning: Learning from Human Feedback without RL
2023/10/20 by Hejna, Joey, Rafailov, Rafael, Sikchi, Harshit +4 · 10 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
2023/02/16 by Harshit Sikchi, Sikchi, Harshit, Amy Zhang +4 · 6 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Modular Robots and Swarm Intelligence #Evolutionary Algorithms and Applications
- f-IRL: Inverse Reinforcement Learning via State Marginal Matching
2020/11/09 by Tianwei Ni, Ni, Tianwei, Harshit Sikchi +9 · 4 citations
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
- SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
2023/11/03 by Sikchi, Harshit, Chitnis, Rohan, Touati, Ahmed +3 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
- Proto Successor Measure: Representing the Behavior Space of an RL Agent
2024/11/29 by Agarwal, Siddhant, Sikchi, Harshit, Stone, Peter +1 · 4 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- CREStE: Scalable Mapless Navigation with Internet Scale Priors and Counterfactual Guidance
2025/03/05 by Zhang, Arthur, Sikchi, Harshit, Zhang, Amy +1 · 5 citations
#Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Robotics (cs.RO)
- A Ranking Game for Imitation Learning
2022/02/07 by Harshit Sikchi, Sikchi, Harshit, Akanksha Saran +5 · 1 citation
Computer Science · #Reinforcement Learning in Robotics #Human Pose and Action Recognition #Multimodal Machine Learning Applications
- RLZero: Direct Policy Inference from Language Without In-Domain Supervision
2024/12/07 by Harshit Sikchi, Sikchi, Harshit, Siddhant Agarwal +15 · 2 voices · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Graphics (cs.GR) #Machine Learning (cs.LG) #Robotics (cs.RO) #cs.AI #cs.GR #cs.LG #cs.RO
- Fast Adaptation with Behavioral Foundation Models
2025/04/10 by Harshit Sikchi, Andrea Tirinzoni, Sikchi, Harshit +14 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics #Robotics (cs.RO)
- A Dual Approach to Imitation Learning from Observations with Offline Datasets
2024/06/13 by Sikchi, Harshit, Chuck, Caleb, Zhang, Amy +1 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)