4 papers
View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification
Quan Zhang, Zeqiang Cai, Peiming Zhao +4
Aerial-Ground Person Re-Identification (AGPReID) remains highly challenging due to drastic viewpoint variations between drones and fixed cameras. Existing methods typically follow…
Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs
Jingze Wu, Quan Zhang, Hongfei Suo +2
Although reinforcement learning (RL) has significantly advanced reasoning capabilities in large multimodal language models (MLLMs), its efficacy remains limited for lightweight mod…
Thinking Before Matching: A Reinforcement Reasoning Paradigm Towards General Person Re-Identification
Quan Zhang, Jingze Wu, Jialong Wang +3
Learning identity-discriminative representations with multi-scene generality has become a critical objective in person re-identification (ReID). However, mainstream perception-driv…
Reinforcing Video Reasoning with Focused Thinking
Jisheng Dang, Jingze Wu, Teng Wang +6
Recent advancements in reinforcement learning, particularly through Group Relative Policy Optimization (GRPO), have significantly improved multimodal large language models for comp…