collaborators

8 papers

cs.RO2026

World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation

Yuhao Pan, Haosong Peng, Zhengshen Zhang +8

Vision-language-action (VLA) models often treat main-view and wrist-view observations as parallel visual inputs, overlooking their distinct roles in robot manipulation. Fine-graine…

cs.AI2026

UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following

Kun Yu, Jianhua Yang, Yixiang Chen +7

The paper introduces UESF-Bench, a large-scale benchmark for unified embodied seeking and following of humans, and presents SeekFollow-VLA, a vision‑language‑action framework that…

cs.CV2026

SpatialBench: Is Your Spatial Foundation Model an All-Round Player?

Haosong Peng, Hao Li, Jiaqi Chen +10

While spatial foundation models have demonstrated impressive performance on standard datasets, a critical question remains: are they truly all-round players capable of generalizing…

cs.CR2026

PrismWF: A Multi-Granularity Patch-Based Transformer for Robust Website Fingerprinting Attack

Yuhao Pan, Wenchao Xu, Fushuo Huo +3

Tor is a low-latency anonymous communication network that protects user privacy by encrypting website traffic. However, recent website fingerprinting (WF) attacks have shown that e…

cs.CR2025

Responsible Diffusion: A Comprehensive Survey on Safety, Ethics, and Trust in Diffusion Models

Kang Wei, Xin Yuan, Fushuo Huo +5

Diffusion models (DMs) have been investigated in various domains due to their ability to generate high-quality data, thereby attracting significant attention. However, similar to t…

cs.CV2025

EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models

Botai Yuan, Yutian Zhou, Yingjie Wang +9

Recent benchmarks for medical Large Vision-Language Models (LVLMs) emphasize leaderboard accuracy, overlooking reliability and safety. We study sycophancy -- models' tendency to un…