From the 1 of 21 linked papers with an AI index.
21 papers
GFR-SAM: Training-Free Referring Camouflaged Object Segmentation via Cross-Image Prompting
Yilong Yang, Jianxin Tian, Shengchuan Zhang +1
The paper introduces GFR‑SAM, a training‑free three‑stage framework that uses cross‑image prompting with SAM3 to segment camouflaged objects referenced by cues, employing exemplar‑…
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding
Shaohui Dai, Yansong Qu, You Shen +2
Recent advances in 3D multimodal large language models (3D-MLLMs) have enabled unified solutions for 3D scene understanding tasks, including visual question answering, captioning,…
Active-SAOOD: Active Sparsely Annotated Oriented Object Detection in Remote Sensing Images
Yu Lin, Jianghang Lin, Kai Ye +2
Reducing the annotation cost of oriented object detection in remote sensing remains a major challenge. Recently, sparse annotation has gained attention for effectively reducing ann…
Discover, Segment, and Select: A Progressive Mechanism for Zero-shot Camouflaged Object Segmentation
Yilong Yang, Jianxin Tian, Shengchuan Zhang +1
Current zero-shot Camouflaged Object Segmentation methods typically employ a two-stage pipeline (discover-then-segment): using MLLMs to obtain visual prompts, followed by SAM segme…
FlashWorld: High-quality 3D Scene Generation within Seconds
Xinyang Li, Tengfei Wang, Zixiao Gu +3
We propose FlashWorld, a generative model that produces 3D scenes from a single image or text prompt in seconds, 10~100 faster than previous works while possessing superior…
SCOUT: Semi-supervised Camouflaged Object Detection by Utilizing Text and Adaptive Data Selection
Weiqi Yan, Lvhai Chen, Shengchuan Zhang +2
The difficulty of pixel-level annotation has significantly hindered the development of the Camouflaged Object Detection (COD) field. To save on annotation costs, previous works lev…