1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2026
MARS: What Retrieval Signals Are Hidden in Multimodal Large Language Models for Text-Video Retrieval?
Uicheol Jung, Juyoung Hong, Geuntaek Lim +1
Text-video retrieval requires representations that can distinguish videos with similar scenes, actions, and temporal patterns. Recent multimodal large language models have been ada…
cs.CV2026★ 1 cited
TAME: Temporal-Aware Mixture-of-Experts for Text-Video Retrieval
Uicheol Jung, Juyoung Hong, Hojung Kwon +1
Text-Video Retrieval (TVR) retrieves videos that match a natural-language query, but extending image-text models such as CLIP to videos is fundamentally limited by the lack of temp…