3 papers
cs.RO2026
Agile-VLA: Few-Shot Industrial Pose Rectification via Implicit Affordance Anchoring
Teng Yan, Zhengyang Pei, Chengyu Shi +8
Deploying Vision-Language-Action (VLA) models on resource-constrained edge platforms encounters a fundamental conflict between high-latency semantic inference and the high-frequenc…
cs.CV2025
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
Xinzhu Li, Juepeng Zheng, Yikun Chen +7
Robust gait recognition requires highly discriminative representations, which are closely tied to input modalities. While binary silhouettes and skeletons have dominated recent lit…
cs.CV2025
DPNet: Dynamic Pooling Network for Tiny Object Detection
Luqi Gong, Haotian Chen, Yikun Chen +4
In unmanned aerial systems, especially in complex environments, accurately detecting tiny objects is crucial. Resizing images is a common strategy to improve detection accuracy, pa…