most citedDiffuVolume: Diffusion Model for Volume based Stereo Matching

6 citations · 7 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV20241 cited

Towards Completeness: A Generalizable Action Proposal Generator for Zero-Shot Temporal Action Localization

Jia-Run Du, Kun-Yu Lin, Jingke Meng +1

To address the zero-shot temporal action localization (ZSTAL) task, existing works develop models that are generalizable to detect and classify actions from unseen categories. They…

cs.CV2024

PRET: Planning with Directed Fidelity Trajectory for Vision and Language Navigation

Renjie Lu, Jingke Meng, Wei-Shi Zheng

Vision and language navigation is a task that requires an agent to navigate according to a natural language instruction. Recent methods predict sub-goals on constructed topology ma…

cs.CV2024

EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding

Yuan-Ming Li, Wei-Jin Huang, An-Lan Wang +3

We present EgoExo-Fitness, a new full-body action understanding dataset, featuring fitness sequence videos recorded from synchronized egocentric and fixed exocentric (third-person)…

cs.CV2024

Rethinking Few-shot Class-incremental Learning: Learning from Yourself

Yu-Ming Tang, Yi-Xing Peng, Jingke Meng +1

Few-shot class-incremental learning (FSCIL) aims to learn sequential classes with limited samples in a few-shot fashion. Inherited from the classical class-incremental learning set…

cs.CV20236 cited

DiffuVolume: Diffusion Model for Volume based Stereo Matching

Dian Zheng, Xiao-Ming Wu, Zuhao Liu +2

Stereo matching is a significant part in many computer vision tasks and driving-based applications. Recently cost volume-based methods have achieved great success benefiting from t…

cs.CV2023

Event-Guided Procedure Planning from Instructional Videos with Text Supervision

An-Lan Wang, Kun-Yu Lin, Jia-Run Du +2

In this work, we focus on the task of procedure planning from instructional videos with text supervision, where a model aims to predict an action sequence to transform the initial…