95 citations · 95 across the 2 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024★ 4 cited
Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases
Liqiong Wang, Teng Jin, Jinyu Yang +3
In the general domain, large multimodal models (LMMs) have achieved significant advancements, yet challenges persist in applying them to specific fields, especially agriculture. As…
cs.CV2023★ 95 cited
Track Anything: Segment Anything Meets Videos
Jinyu Yang, Mingqi Gao, Zhe Li +3
Recently, the Segment Anything Model (SAM) gains lots of attention rapidly due to its impressive segmentation performance on images. Regarding its strong ability on image segmentat…
cs.CV2022
Prompting for Multi-Modal Tracking
Jinyu Yang, Zhe Li, Feng Zheng +2
Multi-modal tracking gains attention due to its ability to be more accurate and robust in complex scenarios compared to traditional RGB-based tracking. Its key lies in how to fuse…