5 papers
Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery
Tingting Chen, Beibei Lin, Srinivas Anumasa +5
Interactive discovery requires agents to maintain and update structured beliefs over many rounds of feedback. Before evaluating agents in noisy, open-ended scientific environments,…
WeatherReasonSeg: A Benchmark for Weather-Aware Reasoning Segmentation in Visual Language Models
Wanjun Du, Zifeng Yuan, Tingting Chen +3
Existing vision-language models (VLMs) have demonstrated impressive performance in reasoning-based segmentation. However, current benchmarks are primarily constructed from high-qua…
RaindropGS: A Benchmark for 3D Gaussian Splatting under Raindrop Conditions
Zhiqiang Teng, Tingting Chen, Beibei Lin +4
3D Gaussian Splatting (3DGS) under raindrop conditions suffers from severe occlusions and optical distortions caused by raindrop contamination on the camera lens, substantially deg…
RGB-to-Polarization Estimation: A New Task and Benchmark Study
Beibei Lin, Zifeng Yuan, Tingting Chen
Polarization images provide rich physical information that is fundamentally absent from standard RGB images, benefiting a wide range of computer vision applications such as reflect…
GeoComplete: Geometry-Aware Diffusion for Reference-Driven Image Completion
Beibei Lin, Tingting Chen, Robby T. Tan
Reference-driven image completion, which restores missing regions in a target view using additional images, is particularly challenging when the target view differs significantly f…