Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
Hongyu Lu, Feng Zhang, Wenwei Jin +5
Large vision-language models (LVLMs) achieve strong multimodal understanding, but their inference cost grows rapidly with the number of visual tokens, especially for high-resolutio…
cs.CV2024
PPTFormer: Pseudo Multi-Perspective Transformer for UAV Segmentation
Deyi Ji, Wenwei Jin, Hongtao Lu +1
The ascension of Unmanned Aerial Vehicles (UAVs) in various fields necessitates effective UAV image segmentation, which faces challenges due to the dynamic perspectives of UAV-capt…
cs.CV2024
Discrete Latent Perspective Learning for Segmentation and Detection
Deyi Ji, Feng Zhao, Lanyun Zhu +3
In this paper, we address the challenge of Perspective-Invariant Learning in machine learning and computer vision, which involves enabling a network to understand images from varyi…