29 citations · 69 across the 22 of their papers we have counts for
11 papers · 1 filter
PhysCaP: Grounding Code-as-Policy Agent with Physics-Informed Exploration
Chen-Yu Lin, Jing-Wen Chen, Hsueh-En Chang +8
We present PhysCaP, a Physics-Informed Code-as-Policy agent for active perception in robotic manipulation. While vision-language-action policies excel at imitating demonstrations,…
Direct Action-Head Injection of A Grounded 3D Point Unlocks Spatial and Task Generalization
Shiang-Feng Tsai, Jin-Cheng Jhang, Yen-Ling Tai +4
Vision-Language-Action (VLA) models leverage large-scale vision-language pretraining for flexible robot manipulation, yet at test time they remain brittle along two axes: spatial g…
HetroD: A High-Fidelity Drone Dataset and Benchmark for Autonomous Driving in Heterogeneous Traffic
Yu-Hsiang Chen, Wei-Jer Chang, Christian Kotulla +7
We present HetroD, a dataset and benchmark for developing autonomous driving systems in heterogeneous environments. HetroD targets the critical challenge of navi- gating real-world…
Affordance-Guided Coarse-to-Fine Exploration for Base Placement in Open-Vocabulary Mobile Manipulation
Tzu-Jung Lin, Jia-Fong Yeh, Hung-Ting Su +3
In open-vocabulary mobile manipulation (OVMM), task success often hinges on the selection of an appropriate base placement for the robot. Existing approaches typically navigate to…
Controllable Collision Scenario Generation via Collision Pattern Prediction
Pin-Lun Chen, Chi-Hsi Kung, Che-Han Chang +2
Evaluating the safety of autonomous vehicles (AVs) requires diverse, safety-critical scenarios, with collisions being especially important yet rare and unsafe to collect in the rea…
GRITS: A Spillage-Aware Guided Diffusion Policy for Robot Food Scooping Tasks
Yen-Ling Tai, Yi-Ru Yang, Kuan-Ting Yu +2
Robotic food scooping is a critical manipulation skill for food preparation and service robots. However, existing robot learning algorithms, especially learn-from-demonstration met…