132 citations · 510 across the 21 of their papers we have counts for
25 papers · 1 filter
LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model
Senqiao Yang, Tianyuan Qu, Xin Lai +4
While LISA effectively bridges the gap between segmentation and large language models to enable reasoning segmentation, it poses certain limitations: unable to distinguish differen…
Prompt Highlighter: Interactive Control for Multi-Modal LLMs
Yuechen Zhang, Shengju Qian, Bohao Peng +2
This study targets a critical aspect of multi-modal LLMs' (LLMs&VLMs) inference: explicit controllable text generation. Multi-modal LLMs empower multi-modality understanding with t…
Stratified Transformer for 3D Point Cloud Segmentation
Xin Lai, Jianhui Liu, Li Jiang +5
3D point cloud segmentation has made tremendous progress in recent years. Most current methods focus on aggregating local features, but fail to directly model long-range dependenci…
SEA: Bridging the Gap Between One- and Two-stage Detector Distillation via SEmantic-aware Alignment
Yixin Chen, Zhuotao Tian, Pengguang Chen +2
We revisit the one- and two-stage detector distillation tasks and present a simple and efficient semantic-aware framework to fill the gap between them. We address the pixel-level i…
Guided Point Contrastive Learning for Semi-supervised Point Cloud Semantic Segmentation
Li Jiang, Shaoshuai Shi, Zhuotao Tian +4
Rapid progress in 3D semantic segmentation is inseparable from the advances of deep network models, which highly rely on large-scale annotated data for training. To address the hig…
Deep Structured Instance Graph for Distilling Object Detectors
Yixin Chen, Pengguang Chen, Shu Liu +2
Effectively structuring deep knowledge plays a pivotal role in transfer from teacher to student, especially in semantic vision tasks. In this paper, we present a simple knowledge s…