3 papers
cs.RO2026
Back to the Familiar Future: Failure Recovery for VLA Policies via Pre-Imagined Milestone Selection
Suyeon Shin, Juwon Kim, Hyeonbin Park +4
Vision-language-action (VLA) policies can deviate from nominal trajectories during manipulation, even when tasks remain physically feasible. Recovering from these deviations is cha…
cs.CV2025
Towards Spatially Consistent Image Generation: On Incorporating Intrinsic Scene Properties into Diffusion Models
Hyundo Lee, Suhyung Choi, Inwoo Hwang +1
Image generation models trained on large datasets can synthesize high-quality images but often produce spatially inconsistent and distorted images due to limited information about…
cs.CV2025
Locality-aware Concept Bottleneck Model
Sujin Jeon, Hyundo Lee, Eungseo Kim +3
Concept bottleneck models (CBMs) are inherently interpretable models that make predictions based on human-understandable visual cues, referred to as concepts. As obtaining dense co…