collaborators

6 papers

cs.CV2026

MHE-Former: Multi-Hypothesis Transformers via Entropy Maximization for 3D Mesh Recovery

Boshu Jia, Rongyu Chen, Linlin Yang +9

Monocular 3D hand and body mesh recovery often suffers from severe occlusion and ambiguity. Traditional deterministic methods typically regress a single optimal solution, leading t…

cs.CL2026

ManGo: Manga Active Narrative Grounding Optimization

Hao Qiu, Junyan Wang, Zheyuan Liu +4

Manga visual question answering requires models to answer questions over panel-based visual narratives, where relevant evidence is distributed across ordered panels, embedded text,…

cs.CV2026

Straight-Path Flow Matching for Incomplete Multi-View Clustering

Yiteng Yuan, Junyan Wang, Zheyuan Liu +4

Incomplete Multi-View Clustering addresses the problem of clustering multi-modal data when certain views are missing. Recent end-to-end generative approaches leverage diffusion mod…

cs.CV2025

EverybodyDance: Bipartite Graph-Based Identity Correspondence for Multi-Character Animation

Haotian Ling, Zequn Chen, Qiuying Chen +6

Consistent pose-driven character animation has achieved remarkable progress in single-character scenarios. However, extending these advances to multi-character settings is non-triv…

cs.CV2025

Cross-Stain Contrastive Learning for Paired Immunohistochemistry and Histopathology Slide Representation Learning

Yizhi Zhang, Lei Fan, Zhulin Tao +4

Universal, transferable whole-slide image (WSI) representations are central to computational pathology. Incorporating multiple markers (e.g., immunohistochemistry, IHC) alongside H…

cs.CV2025

ADNet: A Large-Scale and Extensible Multi-Domain Benchmark for Anomaly Detection Across 380 Real-World Categories

Hai Ling, Jia Guo, Zhulin Tao +6

Anomaly detection (AD) aims to identify defects using normal-only training data. Existing anomaly detection benchmarks (e.g., MVTec-AD with 15 categories) cover only a narrow range…