From the 1 of 11 linked papers with an AI index.
11 papers
PokeNet: Learning Kinematic Models of Articulated Objects from Human Observations
Anmol Gupta, Weiwei Gu, Omkar Patil +2
PokeNet is an end-to-end system that learns the kinematic models of unknown articulated objects from a single human demonstration, predicting joint parameters, manipulation order,…
StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models
Kartikay Milind Pangaonkar, Prabin Rath, Omkar Patil +1
Large scale pre-training on text and image data along with diverse robot demonstrations has helped Vision Language Action models (VLAs) to generalize to novel tasks, objects and sc…
You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector
Omkar Patil, Ondrej Biza, Thomas Weng +9
What happens when a pretrained generative robot policy is provided a constant initial noise as input, rather than repeatedly sampling it from a Gaussian? We demonstrate that the pe…
Spec2Cov: An Agentic Framework for Code Coverage Closure of Digital Hardware Designs
Sean Lowe, Elias Hilaneh, Alma Babbit +3
Hardware verification is one of the most challenging stages of the hardware design process, requiring significant time and resources to ensure a design is fully validated and produ…
Continual Robot Skill and Task Learning via Dialogue
Weiwei Gu, Suresh Kondepudi, Anmol Gupta +2
Interactive robot learning is a challenging problem as the robot is present with human users who expect the robot to learn novel skills to solve novel tasks perpetually with sample…
Meanings and Measurements: Multi-Agent Probabilistic Grounding for Vision-Language Navigation
Swagat Padhan, Lakshya Jain, Bhavya Minesh Shah +3
Robots collaborating with humans must convert natural language goals into actionable, physically grounded decisions. For example, executing a command such as "go two meters to the…