4 papers
Making Single-Cell Data Distillation Auditable: Traceable Real-Cell Coresets via Discrete Min--Max Selection
Yaodi Luo, Peize He, Lingbei Meng +5
Large single-cell datasets are expensive to store, curate, and repeatedly reuse for model training. Data distillation can reduce this burden by building smaller training sets. Howe…
WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents
Bingnan Liu, Chenhang Cui, Rui Huang +7
We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous research-and-engineering by…
VidNum: Diagnosing VLM Failure Modes in Video-Grounded Numerical Reasoning
Shaoyang Cui, Lingbei Meng, Yaodi Luo +1
Video-grounded numerical reasoning requires Vision-Language Models (VLMs) to identify, track, and combine quantitative evidence across frames, actions, and scene changes. Existing…
AugGS: Self-augmented Gaussians with Structural Masks for Sparse-view 3D Reconstruction
Bi'an Du, Lingbei Meng, Wei Hu
Sparse-view 3D reconstruction is a major challenge in computer vision, aiming to create complete three-dimensional models from limited viewing angles. Key obstacles include: 1) a s…