2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 2 cited
Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization
Richard Luo, Austin Peng, Adithya Vasudev +1
Video is an increasingly prominent and information-dense medium, yet it poses substantial challenges for language models. A typical video consists of a sequence of shorter segments…
cs.CV2023
Joint Moment Retrieval and Highlight Detection Via Natural Language Queries
Richard Luo, Austin Peng, Heidi Yap +1
Video summarization has become an increasingly important task in the field of computer vision due to the vast amount of video content available on the internet. In this project, we…