2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Det-SAM2:Technical Report on the Self-Prompting Segmentation Framework Based on Segment Anything Model 2
Zhiting Wang, Qiangong Zhou, Zongyang Liu
Segment Anything Model 2 (SAM2) demonstrates exceptional performance in video segmentation and refinement of segmentation results. We anticipate that it can further evolve to achie…
cs.CL2024★ 2 cited
MathLearner: A Large Language Model Agent Framework for Learning to Solve Mathematical Problems
Wenbei Xie, Donglin Liu, Haoran Yan +2
With the development of artificial intelligence (AI), large language models (LLM) are widely used in many fields. However, the reasoning ability of LLM is still very limited when i…
cs.CV2024
HiLight: Technical Report on the Motern AI Video Language Model
Zhiting Wang, Qiangong Zhou, Kangjie Yang +2
This technical report presents the implementation of a state-of-the-art video encoder for video-text modal alignment and a video conversation framework called HiLight, which featur…