works on

From the 1 of 8 linked papers with an AI index.

collaborators

8 papers

cs.CV2026

Semantic-Aware Temporal Adaptation for UAV Anti-UAV Tracking

Xiaozhen Qiao, Da Zhang, Yubin Guo +3

The paper introduces SATATrack, a framework that uses semantic descriptions of the target and online feature distribution alignment to improve UAV-to-UAV tracking under rapid appea…

cs.CV2026

AHAP: Reconstructing Arbitrary Humans from Arbitrary Perspectives with Geometric Priors

Xiaozhen Qiao, Wenjia Wang, Zhiyuan Zhao +4

Reconstructing 3D humans from images captured at multiple perspectives typically requires pre-calibration, like using checkerboards or MVS algorithms, which limits scalability and…

cs.CV2026

EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos

Hongming Fu, Wenjia Wang, Xiaozhen Qiao +5

We propose EgoGrasp, the first method to reconstruct world-space hand-object interactions (W-HOI) from dynamic egoview videos, supporting open-vocabulary objects. Accurate W-HOI re…

cs.CV2026

Mitigating Long-Tail Bias in HOI Detection via Adaptive Diversity Cache

Yuqiu Jiang, Xiaozhen Qiao, Yifan Chen +3

Human-Object Interaction (HOI) detection is a fundamental task in computer vision, empowering machines to comprehend human-object relationships in diverse real-world scenarios. Rec…

cs.RO2026

Bridging Perception and Planning: Towards End-to-End Planning for Signal Temporal Logic Tasks

Bowen Ye, Junyue Huang, Yang Liu +2

We investigate the task and motion planning problem for Signal Temporal Logic (STL) specifications in robotics. Existing STL methods rely on pre-defined maps or mobility representa…

cs.CV2025

DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation

Xiangchen Yin, Jiahui Yuan, Zhangchi Hu +5

Existing video Variational Autoencoders (VAEs) generally overlook the similarity between frame contents, leading to redundant latent modeling. In this paper, we propose decoupled V…