1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.AI2025★ 1 cited
T2UE: Generating Unlearnable Examples from Text Descriptions
Xingjun Ma, Hanxun Huang, Tianwei Song +3
Large-scale pre-training frameworks like CLIP have revolutionized multimodal learning, but their reliance on web-scraped datasets, frequently containing private user data, raises s…
cs.CV2025
SAMA: Towards Multi-Turn Referential Grounded Video Chat with Large Language Models
Ye Sun, Hao Zhang, Henghui Ding +3
Achieving fine-grained spatio-temporal understanding in videos remains a major challenge for current Video Large Multimodal Models (Video LMMs). Addressing this challenge requires…