5 papers
Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning
Matthew Foutter, Matteo Cercola, Lena Wild +4
Embodied Chain-of-Thought has emerged as a promising mechanism to enhance robot decision-making and interpretability in black-box Vision-Language Action (VLA) models. However, whet…
Bridging Structure and Language: Graph-Based Visual Reasoning for Autonomous Road Understanding
Lena Wild, Katie Z Luo, Marco Pavone
Structured road understanding of lane geometry, topology, and traffic element relationships is foundational to safe autonomous driving. While vision-language models (VLMs) offer pr…
ArgoTweak: Towards Self-Updating HD Maps through Structured Priors
Lena Wild, Rafael Valencia, Patric Jensfelt
Reliable integration of prior information is crucial for self-verifying and self-updating HD maps. However, no public dataset includes the required triplet of prior maps, current m…
ExelMap: Explainable Element-based HD-Map Change Detection and Update
Lena Wild, Ludvig Ericson, Rafael Valencia +1
Acquisition and maintenance are central problems in deploying high-definition (HD) maps for autonomous driving, with two lines of research prevalent in current literature: Online H…
Coherent Perfect Absorption of Arbitrary Wavefronts at an Exceptional Point
Helmut Hörner, Lena Wild, Yevgeny Slobodkin +3
A Coherent Perfect Absorber (CPA) exploits the interferometric nature of light to deposit all of a light field's incident energy into an otherwise weakly absorbing sample. The down…