most citedVTG-GPT: Tuning-Free Zero-Shot Video Temporal Grounding with GPT

19 citations · 49 across the 5 of their papers we have counts for

collaborators

5 papers

quant-ph2024

Quantum state transfer between superconducting cavities via exchange-free interactions

Jie Zhou, Ming Li, Weiting Wang +10

We propose and experimentally demonstrate a novel protocol for transferring quantum states between superconducting cavities using only continuous two-mode squeezing interactions, w…

cs.CV202412 cited

GPTSee: Enhancing Moment Retrieval and Highlight Detection via Description-Based Similarity Features

Yunzhuo Sun, Yifang Xu, Zien Xie +2

Moment retrieval (MR) and highlight detection (HD) aim to identify relevant moments and highlights in video from corresponding natural language query. Large language models (LLMs)…

cs.CV202419 cited

VTG-GPT: Tuning-Free Zero-Shot Video Temporal Grounding with GPT

Yifang Xu, Yunzhuo Sun, Zien Xie +2

Video temporal grounding (VTG) aims to locate specific temporal segments from an untrimmed video based on a linguistic query. Most existing VTG models are trained on extensive anno…

cs.CV202414 cited

Pyramid Feature Attention Network for Monocular Depth Prediction

Yifang Xu, Chenglei Peng, Ming Li +2

Deep convolutional neural networks (DCNNs) have achieved great success in monocular depth estimation (MDE). However, few existing works take the contributions for MDE of different…

cs.CV20234 cited

MH-DETR: Video Moment and Highlight Detection with Cross-modal Transformer

Yifang Xu, Yunzhuo Sun, Yang Li +3

With the increasing demand for video understanding, video moment and highlight detection (MHD) has emerged as a critical research topic. MHD aims to localize all moments and predic…