works on

From the 2 of 16 linked papers with an AI index.

collaborators

16 papers

cs.CV2026

MedXplore: Towards Reliable and Unbiased Generalized Category Discovery in Medical Imaging

Jianwei He, Kailin Lyu, Junhao Dong +6

The paper presents MedXplore, a unified framework for generalized category discovery in medical imaging that leverages frequency-domain adaptive attention and an adaptive cosine-an…

cs.CV2026

VQ-Touch: A Data-Efficient Tactile Generation Framework Across Sensors and Scenarios

Kailin Lyu, Long Xiao, Jianing Zeng +3

The paper presents VQ-Touch, a framework that efficiently generates high‑fidelity tactile images across different sensors and scenarios using a VQ‑GAN based representation and a di…

cs.AI2026

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

Kailin Lyu, Di Wu, Pengwei Zhang +12

Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language systems for tactile commonsense re…

cs.AI2026

TacReasoner: A Dynamic Tactile-Language Framework for Interactive Reasoning in Real-World Scenarios

Kailin Lyu, Di Wu, Long Xiao +7

Among the five primary human senses, tactile is arguably the most fundamental to survival, as it enables the perception of physical contact and interaction in real-world environmen…

cs.CV2026

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

Kangyi Wu, Pengna Li, Kailin Lyu +5

Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Large Language Models(Video-LLM…

cs.AI2026

PAL-Bench: Evidence-Grounded Profile Reconstruction from Longitudinal Personal Albums

Qiwei Yan, Zhiqiang Yuan, Zexi Jia +4

Longitudinal personal albums are weak-schema multimodal databases: noisy perceptual records whose key facts require joins across faces, text, timestamps, locations, and repeated ev…