activity
20242026
collaborators

6 papers

cs.CV2026

ResemBrick: Brick Reconstruction from Photographs with Perceptual Fidelity and Buildability

Xilun Chen, Hanwen Wan, Yusong Zhao +3

Producing a hand-buildable, colored brick model of a 3D object from a few casual photographs is a clean testbed for a broader challenge: generating 3D content that meets hard physi…

cs.CV2026

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

Hanwen Wan, Zexin Lin, Yixuan Deng +1

Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However, existing benchmarks primarily focus…

cs.RO2025

A Learning-based Control Methodology for Transitioning VTOL UAVs

Zexin Lin, Yebin Zhong, Hanwen Wan +3

Transition control poses a critical challenge in Vertical Take-Off and Landing Unmanned Aerial Vehicle (VTOL UAV) development due to the tilting rotor mechanism, which shifts the c…

cs.RO2025

EmbodiedAgent: A Scalable Hierarchical Approach to Overcome Practical Challenge in Multi-Robot Control

Hanwen Wan, Yifei Chen, Yixuan Deng +6

This paper introduces EmbodiedAgent, a hierarchical framework for heterogeneous multi-robot control. EmbodiedAgent addresses critical limitations of hallucination in impractical ta…

cs.RO2025

GenTe: Generative Real-world Terrains for General Legged Robot Locomotion Control

Hanwen Wan, Mengkang Li, Donghao Wu +4

Developing bipedal robots capable of traversing diverse real-world terrains presents a fundamental robotics challenge, as existing methods using predefined height maps and static e…

cs.CL2024

ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models

Yuxiang Zhang, Jing Chen, Junjie Wang +10

Tool-augmented large language models (LLMs) are rapidly being integrated into real-world applications. Due to the lack of benchmarks, the community has yet to fully understand the…