1 citations · 3 across the 6 of their papers we have counts for
3 papers
cs.CV2026
ARM: An AutoRegressive Large Multimodal Model with Unified Discrete Representations
Junke Wang, Xiao Wang, Jiacheng Pan +16
This paper introduces ARM, a discrete representation-based AutoRegressive Model that unifies image understanding, generation, and editing within a next-token prediction framework.…
cs.CV2022★ 1 cited
Renmin University of China at TRECVID 2022: Improving Video Search by Feature Fusion and Negation Understanding
Xirong Li, Aozhu Chen, Ziyue Wang +4
We summarize our TRECVID 2022 Ad-hoc Video Search (AVS) experiments. Our solution is built with two new techniques, namely Lightweight Attentional Feature Fusion (LAFF) for combini…
cs.CV2021★ 1 cited
Unsupervised Domain Expansion for Visual Categorization
Jie Wang, Kaibin Tian, Dayong Ding +2
Expanding visual categorization into a novel domain without the need of extra annotation has been a long-term interest for multimedia intelligence. Previously, this challenge has b…