20 citations · 24 across the 6 of their papers we have counts for
6 papers
ZEETAD: Adapting Pretrained Vision-Language Model for Zero-Shot End-to-End Temporal Action Detection
Thinh Phan, Khoa Vo, Duy Le +3
Temporal action detection (TAD) involves the localization and classification of action instances within untrimmed videos. While standard TAD follows fully supervised learning with…
Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation
Kashu Yamazaki, Taisei Hanyu, Khoa Vo +5
Precise 3D environmental mapping is pivotal in robotics. Existing methods often rely on predefined concepts during training or are time-intensive when generating semantic maps. Thi…
Distributionally Robust Cross Subject EEG Decoding
Tiehang Duan, Zhenyi Wang, Gianfranco Doretto +3
Recently, deep learning has shown to be effective for Electroencephalography (EEG) decoding tasks. Yet, its performance can be negatively influenced by two key factors: 1) the high…
ChatGPT in the Age of Generative AI and Large Language Models: A Concise Survey
Salman Mohamadi, Ghulam Mujtaba, Ngan Le +2
ChatGPT is a large language model (LLM) created by OpenAI that has been carefully trained on a large amount of data. It has revolutionized the field of natural language processing…
A Robust Likelihood Model for Novelty Detection
Ranya Almohsen, Shivang Patel, Donald A. Adjeroh +1
Current approaches to novelty or anomaly detection are based on deep neural networks. Despite their effectiveness, neural networks are also vulnerable to imperceptible deformations…
Self-supervised Interest Point Detection and Description for Fisheye and Perspective Images
Marcela Mera-Trujillo, Shivang Patel, Yu Gu +1
Keypoint detection and matching is a fundamental task in many computer vision problems, from shape reconstruction, to structure from motion, to AR/VR applications and robotics. It…