Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking
Weiguang Zhao, Haoran Xu, Xingyu Miao +11
Point tracking aims to follow visual points through complex motion, occlusion, and viewpoint changes, and has advanced rapidly with modern foundation models. Yet progress toward ge…
cs.CV2024
MammothModa: Multi-Modal Large Language Model
Qi She, Junwen Pan, Xin Wan +3
In this report, we introduce MammothModa, yet another multi-modal large language model (MLLM) designed to achieve state-of-the-art performance starting from an elementary baseline.…