6 papers
ResemBrick: Brick Reconstruction from Photographs with Perceptual Fidelity and Buildability
Xilun Chen, Hanwen Wan, Yusong Zhao +3
Producing a hand-buildable, colored brick model of a 3D object from a few casual photographs is a clean testbed for a broader challenge: generating 3D content that meets hard physi…
Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness
Hanwen Wan, Zexin Lin, Yixuan Deng +1
Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However, existing benchmarks primarily focus…
A Learning-based Control Methodology for Transitioning VTOL UAVs
Zexin Lin, Yebin Zhong, Hanwen Wan +3
Transition control poses a critical challenge in Vertical Take-Off and Landing Unmanned Aerial Vehicle (VTOL UAV) development due to the tilting rotor mechanism, which shifts the c…
EmbodiedAgent: A Scalable Hierarchical Approach to Overcome Practical Challenge in Multi-Robot Control
Hanwen Wan, Yifei Chen, Yixuan Deng +6
This paper introduces EmbodiedAgent, a hierarchical framework for heterogeneous multi-robot control. EmbodiedAgent addresses critical limitations of hallucination in impractical ta…
GenTe: Generative Real-world Terrains for General Legged Robot Locomotion Control
Hanwen Wan, Mengkang Li, Donghao Wu +4
Developing bipedal robots capable of traversing diverse real-world terrains presents a fundamental robotics challenge, as existing methods using predefined height maps and static e…
ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models
Yuxiang Zhang, Jing Chen, Junjie Wang +10
Tool-augmented large language models (LLMs) are rapidly being integrated into real-world applications. Due to the lack of benchmarks, the community has yet to fully understand the…