10 papers
Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case
Shilin Hu, Jingyi Xu, Dimitris Samaras +1
Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural question: do they reduce the…
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
Shilin Hu, Jingyi Xu, Sagnik Das +2
Shadows encode rich information about scene geometry and illumination, yet existing methods either predict a unified shadow mask or overlook attached shadows entirely. We address t…
Counting Trees from Satellite Imagery with Noisy Supervision
Dimitri Gominski, Maurice Mugabowindekwe, Qiue Xu +6
Counting individual trees is a fundamental task for environmental monitoring, yet remains largely unexplored with satellite imagery. At these resolutions, isolated trees may still…
Phrase-Instance Alignment for Generalized Referring Segmentation
E-Ro Nguyen, Hieu Le, Dimitris Samaras +1
Generalized Referring expressions can describe one object, several related objects, or none at all. Existing generalized referring segmentation (GRES) models treat all cases alike,…
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
Shilin Hu, Jingyi Xu, Akshat Dave +2
Generating realistic shadows for inserted objects requires reasoning about scene geometry and illumination. However, most existing methods operate purely in image space, leaving th…
Personalized Image Descriptions from Attention Sequences
Ruoyu Xue, Hieu Le, Jingyi Xu +5
People can view the same image differently: they focus on different regions, objects, and details in varying orders and describe them in distinct linguistic styles. This leads to s…