1 citations · 1 across the 4 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024★ 1 cited
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
Hao Liang, Jiapeng Li, Tianyi Bai +7
Recently, with the rise of web videos, managing and understanding large-scale video datasets has become increasingly important. Video Large Language Models (VideoLLMs) have emerged…
cs.CV2024
Are Bigger Encoders Always Better in Vision Large Models?
Bozhou Li, Hao Liang, Zimo Meng +1
In recent years, multimodal large language models (MLLMs) have shown strong potential in real-world applications. They are developing rapidly due to their remarkable ability to com…