Mehta, Viraj
- Group Robust Preference Optimization in Reward-free RLHF
2024/05/30 by Shyam Sundhar Ramesh, Ramesh, Shyam Sundhar, Yifan Hu +11 · 20 citations
Computer Science · #Advanced Multi-Objective Optimization Algorithms
- Learning Task-Oriented Grasping for Tool Manipulation from Simulated Self-Supervision
2018/06/25 by Kuan Fang, Yuke Zhu, Fang, Kuan +11 · 6 citations
Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Muscle activation and electromyography studies #Robot Manipulation and Learning #Robotics (cs.RO) #Soft Robotics and Applications
- Representational aspects of depth and conditioning in normalizing flows
2020/10/02 by Frederic Koehler, Viraj Mehta, Koehler, Frederic +3 · 4 citations
Computer Science · Medicine · #Generative Adversarial Networks and Image Synthesis #Computer Graphics and Visualization Techniques #Advanced Neuroimaging Techniques and Applications
- Sample Efficient Preference Alignment in LLMs via Active Exploration
2023/12/01 by Mehta, Viraj, Belakaria, Syrine, Das, Vikramjeet +7 · 5 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- DeformNet: Free-Form Deformation Network for 3D Shape Reconstruction from a Single Image
2017/08/11 by Kurenkov, Andrey, Ji, Jingwei, Garg, Animesh +4 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Graphics (cs.GR)
- Variational autoencoders in the presence of low-dimensional data: landscape and implicit bias
2021/12/13 by Koehler, Frederic, Mehta, Viraj, Zhou, Chenghui +1 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Exploration via Planning for Information about the Optimal Trajectory
2022/10/06 by Viraj Mehta, Ian Char, Mehta, Viraj +13 · 1 citation
Computer Science · #Reinforcement Learning in Robotics #Machine Learning and Data Classification #Data Stream Mining Techniques
- Near-optimal Policy Identification in Active Reinforcement Learning
2022/12/19 by Xiang Li, Viraj Mehta, Li, Xiang +13 · 1 citation
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Machine Learning and Algorithms