4 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
Bo He, Hengduo Li, Young Kyun Jang +5
With the success of large language models (LLMs), integrating the vision model into LLMs to build vision-language foundation models has gained much more interest recently. However,…
cs.CL2022★ 2 cited
Searching for Structure in Unfalsifiable Claims
Peter Ebert Christensen, Frederik Warburg, Menglin Jia +1
Social media platforms give rise to an abundance of posts and comments on every topic imaginable. Many of these posts express opinions on various aspects of society, but their unfa…
cs.CV2021★ 4 cited
Rethinking Nearest Neighbors for Visual Classification
Menglin Jia, Bor-Chun Chen, Zuxuan Wu +3
Neural network classifiers have become the de-facto choice for current "pre-train then fine-tune" paradigms of visual classification. In this paper, we investigate k-Nearest-Neighb…