collaborators

14 papers

cs.CV2026

Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case

Shilin Hu, Jingyi Xu, Dimitris Samaras +1

Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural question: do they reduce the…

cs.CV2026

Learning to Generate Multiple Objects from Dense and Occluded Layouts

Bach-Hoang Ngo, Si-Tri Ngo, Hieu Le +1

Text-to-image diffusion models fail to generate correct object counts in dense scenes, where overlapping instances collapse into indistinguishable structures despite appearing visu…

cs.CV2026

Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning

Shilin Hu, Jingyi Xu, Sagnik Das +2

Shadows encode rich information about scene geometry and illumination, yet existing methods either predict a unified shadow mask or overlook attached shadows entirely. We address t…

cs.CV2026

Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer

Haoyu Wu, Jingyi Xu, Qiaomu Miao +2

Rotary positional embeddings (RoPE) are widely used in diffusion transformers (DiTs) to encode spatial relationships, yet their behavior with mixed-resolution tokens remains undere…

cs.CV2026

Counting Trees from Satellite Imagery with Noisy Supervision

Dimitri Gominski, Maurice Mugabowindekwe, Qiue Xu +6

Counting individual trees is a fundamental task for environmental monitoring, yet remains largely unexplored with satellite imagery. At these resolutions, isolated trees may still…

cs.CV2026

Phrase-Instance Alignment for Generalized Referring Segmentation

E-Ro Nguyen, Hieu Le, Dimitris Samaras +1

Generalized Referring expressions can describe one object, several related objects, or none at all. Existing generalized referring segmentation (GRES) models treat all cases alike,…