activity
20192022
most citedA Spatio-temporal Attention-based Model for Infant Movement Assessment from Videos

53 citations · 79 across the 9 of their papers we have counts for

collaborators

18 papers

cs.CV2022

Guiding Visual Question Answering with Attention Priors

Thao Minh Le, Vuong Le, Sunil Gupta +2

The current success of modern visual reasoning systems is arguably attributed to cross-modality attention mechanisms. However, in deliberative reasoning such as in VQA, attention i…

cs.CV2022

Persistent-Transient Duality in Human Behavior Modeling

Hung Tran, Vuong Le, Svetha Venkatesh +1

We propose to model the persistent-transient duality in human behavior using a parent-child multi-channel neural network, which features a parent persistent channel that manages th…

cs.LG20211 cited

A Field Guide to Scientific XAI: Transparent and Interpretable Deep Learning for Bioinformatics Research

Thomas P Quinn, Sunil Gupta, Svetha Venkatesh +1

Deep learning has become popular because of its potential to achieve high accuracy in prediction tasks. However, accuracy is not always the only goal of statistical modelling, espe…

cs.CV20216 cited

Hierarchical Object-oriented Spatio-Temporal Reasoning for Video Question Answering

Long Hoang Dang, Thao Minh Le, Vuong Le +1

Video Question Answering (Video QA) is a powerful testbed to develop new AI capabilities. This task necessitates learning to reason about objects, relations, and events across visu…

cs.CV202153 cited

A Spatio-temporal Attention-based Model for Infant Movement Assessment from Videos

Binh Nguyen-Thai, Vuong Le, Catherine Morgan +3

The absence or abnormality of fidgety movements of joints or limbs is strongly indicative of cerebral palsy in infants. Developing computer-based methods for assessing infant movem…

cs.CV2021

Object-Centric Representation Learning for Video Question Answering

Long Hoang Dang, Thao Minh Le, Vuong Le +1

Video question answering (Video QA) presents a powerful testbed for human-like intelligent behaviors. The task demands new capabilities to integrate video processing, language unde…