Vineet Bhat
- 3D CAVLA: Leveraging Depth and 3D Context to Generalize Vision Language Action Models for Unseen Tasks
2025/05/09 by Vineet Bhat, Bhat, Vineet, Lan, Yu-Hsiang +6 · 8 citations
Computer Science · Engineering · #Multimodal Machine Learning Applications #Robot Manipulation and Learning #Advanced Neural Network Applications
- HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
2024/09/16 by Vineet Bhat, P. Krishnamurthy, Bhat, Vineet +5 · 2 citations
Computer Science · #Multimodal Machine Learning Applications #Hand Gesture Recognition Systems #Advanced Image and Video Retrieval Techniques