vix.ing · top · new · best · stats · spec

Sikchi, Harshit

  1. gpt-oss-120b & gpt-oss-20b Model Card
    2025/08/08 by OpenAI, :, Agarwal, Sandhini +124 · 251 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  2. Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
    2024/06/05 by Rafael Rafailov, Yaswanth Chittepu, Rafailov, Rafael +13 · 16 citations
    Engineering · Decision Sciences · Computer Science · #Advanced Control Systems Optimization #Simulation Techniques and Applications #Statistical and Computational Modeling
  3. Learning Off-Policy with Online Planning
    2020/08/23 by Sikchi, Harshit, Zhou, Wenxuan, Held, David · 6 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  4. Contrastive Preference Learning: Learning from Human Feedback without RL
    2023/10/20 by Hejna, Joey, Rafailov, Rafael, Sikchi, Harshit +4 · 10 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  5. Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
    2023/02/16 by Harshit Sikchi, Sikchi, Harshit, Amy Zhang +4 · 6 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Modular Robots and Swarm Intelligence #Evolutionary Algorithms and Applications
  6. f-IRL: Inverse Reinforcement Learning via State Marginal Matching
    2020/11/09 by Tianwei Ni, Ni, Tianwei, Harshit Sikchi +9 · 4 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  7. SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
    2023/11/03 by Sikchi, Harshit, Chitnis, Rohan, Touati, Ahmed +3 · 3 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  8. Proto Successor Measure: Representing the Behavior Space of an RL Agent
    2024/11/29 by Agarwal, Siddhant, Sikchi, Harshit, Stone, Peter +1 · 4 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  9. CREStE: Scalable Mapless Navigation with Internet Scale Priors and Counterfactual Guidance
    2025/03/05 by Zhang, Arthur, Sikchi, Harshit, Zhang, Amy +1 · 5 citations
    #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Robotics (cs.RO)
  10. A Ranking Game for Imitation Learning
    2022/02/07 by Harshit Sikchi, Sikchi, Harshit, Akanksha Saran +5 · 1 citation
    Computer Science · #Reinforcement Learning in Robotics #Human Pose and Action Recognition #Multimodal Machine Learning Applications
  11. RLZero: Direct Policy Inference from Language Without In-Domain Supervision
    2024/12/07 by Harshit Sikchi, Sikchi, Harshit, Siddhant Agarwal +15 · 2 voices · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Graphics (cs.GR) #Machine Learning (cs.LG) #Robotics (cs.RO) #cs.AI #cs.GR #cs.LG #cs.RO
  12. Fast Adaptation with Behavioral Foundation Models
    2025/04/10 by Harshit Sikchi, Andrea Tirinzoni, Sikchi, Harshit +14 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics #Robotics (cs.RO)
  13. A Dual Approach to Imitation Learning from Observations with Offline Datasets
    2024/06/13 by Sikchi, Harshit, Chuck, Caleb, Zhang, Amy +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)