activity
20242026
collaborators

8 papers

cs.CV2026

Beyond Thinking: Imagining in 360 for Humanoid Visual Search

Jingdong Zhang, Yizhou Wang, Zhengzhong Tu +3

Humanoid Visual Search (HVS) requires agents to actively explore immersive 360 environments. While prior methods treat this as a monolithic task relying on cumulative, mult…

cs.CV2026

MTPano: Multi-Task Panoramic Scene Understanding via Label-Free Integration of Dense Prediction Priors

Jingdong Zhang, Xiaohang Zhan, Lingzhi Zhang +5

Comprehensive panoramic scene understanding is critical for immersive applications, yet it remains challenging due to the scarcity of high-resolution, multi-task annotations. While…

cs.CV2026

UniSER: A Foundation Model for Unified Soft Effects Removal

Jingdong Zhang, Lingzhi Zhang, Qing Liu +12

Digital images are often degraded by soft effects such as lens flare, haze, shadows, and reflections, which reduce aesthetics even though the underlying pixels remain partially vis…

cs.CV2025

SPGen: Spherical Projection as Consistent and Flexible Representation for Single Image 3D Shape Generation

Jingdong Zhang, Weikai Chen, Yuan Liu +6

Existing single-view 3D generative models typically adopt multiview diffusion priors to reconstruct object surfaces, yet they remain prone to inter-view inconsistencies and are una…

cs.CV2025

Multi-Task Label Discovery via Hierarchical Task Tokens for Partially Annotated Dense Predictions

Jingdong Zhang, Hanrong Ye, Xin Li +2

In recent years, simultaneous learning of multiple dense prediction tasks with partially annotated label data has emerged as an important research area. Previous works primarily fo…

cs.CV2025

3R-GS: Best Practice in Optimizing Camera Poses Along with 3DGS

Zhisheng Huang, Peng Wang, Jingdong Zhang +3

3D Gaussian Splatting (3DGS) has revolutionized neural rendering with its efficiency and quality, but like many novel view synthesis methods, it heavily depends on accurate camera…