Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Det-SAM2:Technical Report on the Self-Prompting Segmentation Framework Based on Segment Anything Model 2
Zhiting Wang, Qiangong Zhou, Zongyang Liu
Segment Anything Model 2 (SAM2) demonstrates exceptional performance in video segmentation and refinement of segmentation results. We anticipate that it can further evolve to achie…
cs.CV2024
HiLight: Technical Report on the Motern AI Video Language Model
Zhiting Wang, Qiangong Zhou, Kangjie Yang +2
This technical report presents the implementation of a state-of-the-art video encoder for video-text modal alignment and a video conversation framework called HiLight, which featur…