2 citations · 2 across the 3 of their papers we have counts for
3 papers
VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation
Haochen Zhang, Nader Zantout, Pujith Kachana +3
With the recent rise of Large Language Models (LLMs), Vision-Language Models (VLMs), and other general foundation models, there is growing potential for multimodal, multi-task embo…
Snail: Secure Single Iteration Localization
James Choncholas, Pujith Kachana, André Mateus +2
Localization is a computer vision task by which the position and orientation of a camera is determined from an image and environmental map. We propose a method for performing local…
Neural Field Dynamics Model for Granular Object Piles Manipulation
Shangjie Xue, Shuo Cheng, Pujith Kachana +1
We present a learning-based dynamics model for granular material manipulation. Inspired by the Eulerian approach commonly used in fluid dynamics, our method adopts a fully convolut…