8 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Generative Human-Object Interaction Detection via Differentiable Cognitive Steering of Multi-modal LLMs
Zhaolin Cai, Huiyu Duan, Zitong Xu +6
Human-object interaction (HOI) detection aims to localize human-object pairs and the interactions between them. Existing methods operate under a closed-world assumption, treating t…
cs.CV2025
SAM-DAQ: Segment Anything Model with Depth-guided Adaptive Queries for RGB-D Video Salient Object Detection
Jia Lin, Xiaofei Zhou, Jiyuan Liu +4
Recently segment anything model (SAM) has attracted widespread concerns, and it is often treated as a vision foundation model for universal segmentation. Some researchers have atte…
cs.CV2024★ 8 cited
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
Pengyuan Zhou, Lin Wang, Zhi Liu +4
This paper offers an insightful examination of how currently top-trending AI technologies, i.e., generative artificial intelligence (Generative AI) and large language models (LLMs)…