4 papers
When W4A4 Breaks Camouflaged Object Detection: Token-Group Dual-Constraint Activation Quantization
Tianqi Li, Wenyu Fang, Xin He +3
The paper proposes a post‑training 4‑bit activation quantization method for transformer‑based camouflaged object detection that mitigates token‑level range domination to preserve s…
STAC: Selective Spatiotemporal Aggregation and Compression for Video Reasoning Segmentation
Syed Ariff Syed Hesham, Yun Liu, Guolei Sun +4
Video reasoning segmentation demands pixel-accurate object tracking across hundreds of frames under complex natural language queries, producing dense spatiotemporal tokens whose qu…
Evaluating SAM2 for Video Semantic Segmentation
Syed Hesham Syed Ariff, Yun Liu, Guolei Sun +4
The Segmentation Anything Model 2 (SAM2) has proven to be a powerful foundation model for promptable visual object segmentation in both images and videos, capable of storing object…
Exploiting Temporal State Space Sharing for Video Semantic Segmentation
Syed Ariff Syed Hesham, Yun Liu, Guolei Sun +5
Video semantic segmentation (VSS) plays a vital role in understanding the temporal evolution of scenes. Traditional methods often segment videos frame-by-frame or in a short tempor…