Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
ResemBrick: Brick Reconstruction from Photographs with Perceptual Fidelity and Buildability
Xilun Chen, Hanwen Wan, Yusong Zhao +3
Producing a hand-buildable, colored brick model of a 3D object from a few casual photographs is a clean testbed for a broader challenge: generating 3D content that meets hard physi…
cs.CV2025
Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness
Hanwen Wan, Zexin Lin, Yixuan Deng +1
Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However, existing benchmarks primarily focus…