8 citations · 14 across the 9 of their papers we have counts for
9 papers · 1 filter
Efficient Human-Contact Representation for Human-Scene Interaction
Nghia Vu, Tuong Do, Binh X. Nguyen +2
Human-scene interaction is an active research topic with several industrial applications in virtual reality, gaming, robotics, and surveillance. Despite significant progress in net…
AffordMatcher: Affordance Learning in 3D Scenes from Visual Signifiers
Nghia Vu, Tuong Do, Khang Nguyen +8
Affordance learning is a complex challenge in many applications, where existing approaches primarily focus on the geometric structures, visual knowledge, and affordance labels of o…
Lightweight Temporal Transformer Decomposition for Federated Autonomous Driving
Tuong Do, Binh X. Nguyen, Quang D. Tran +3
Traditional vision-based autonomous driving systems often face difficulties in navigating complex environments when relying solely on single-image inputs. To overcome this limitati…
Coarse-to-Fine Reasoning for Visual Question Answering
Binh X. Nguyen, Tuong Do, Huy Tran +3
Bridging the semantic gap between image and question is an important step to improve the accuracy of the Visual Question Answering (VQA) task. However, most of the existing VQA met…
Multiple Meta-model Quantifying for Medical Visual Question Answering
Tuong Do, Binh X. Nguyen, Erman Tjiputra +3
Transfer learning is an important step to extract meaningful features and overcome the data limitation in the medical Visual Question Answering (VQA) task. However, most of the exi…
Graph-based Person Signature for Person Re-Identifications
Binh X. Nguyen, Binh D. Nguyen, Tuong Do +3
The task of person re-identification (ReID) is to match images of the same person over multiple non-overlapping camera views. Due to the variations in visual factors, previous work…