activity
20192026
most citedDeep Metric Learning Meets Deep Clustering: An Novel Unsupervised Approach for Feature Embedding

8 citations · 14 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

9 papers · 1 filter

cs.CV2026

Efficient Human-Contact Representation for Human-Scene Interaction

Nghia Vu, Tuong Do, Binh X. Nguyen +2

Human-scene interaction is an active research topic with several industrial applications in virtual reality, gaming, robotics, and surveillance. Despite significant progress in net…

cs.CV2026

AffordMatcher: Affordance Learning in 3D Scenes from Visual Signifiers

Nghia Vu, Tuong Do, Khang Nguyen +8

Affordance learning is a complex challenge in many applications, where existing approaches primarily focus on the geometric structures, visual knowledge, and affordance labels of o…

cs.CV2025

Lightweight Temporal Transformer Decomposition for Federated Autonomous Driving

Tuong Do, Binh X. Nguyen, Quang D. Tran +3

Traditional vision-based autonomous driving systems often face difficulties in navigating complex environments when relying solely on single-image inputs. To overcome this limitati…

cs.CV2021

Coarse-to-Fine Reasoning for Visual Question Answering

Binh X. Nguyen, Tuong Do, Huy Tran +3

Bridging the semantic gap between image and question is an important step to improve the accuracy of the Visual Question Answering (VQA) task. However, most of the existing VQA met…

cs.CV2021★ 5 cited

Multiple Meta-model Quantifying for Medical Visual Question Answering

Tuong Do, Binh X. Nguyen, Erman Tjiputra +3

Transfer learning is an important step to extract meaningful features and overcome the data limitation in the medical Visual Question Answering (VQA) task. However, most of the exi…

cs.CV2021

Graph-based Person Signature for Person Re-Identifications

Binh X. Nguyen, Binh D. Nguyen, Tuong Do +3

The task of person re-identification (ReID) is to match images of the same person over multiple non-overlapping camera views. Due to the variations in visual factors, previous work…