19 citations · 19 across the 1 of their papers we have counts for
3 papers
cs.CV2024
Are Visual-Language Models Effective in Action Recognition? A Comparative Study
Mahmoud Ali, Di Yang, François Brémond
Current vision-language foundation models, such as CLIP, have recently shown significant improvement in performance across various downstream tasks. However, whether such foundatio…
cs.CV2023★ 19 cited
MultiMediate'23: Engagement Estimation and Bodily Behaviour Recognition in Social Interactions
Philipp Müller, Michal Balazia, Tobias Baur +9
Automatic analysis of human behaviour is a fundamental prerequisite for the creation of machines that can effectively interact with- and support humans in social interactions. In M…
cs.CV2022
Multimodal Vision Transformers with Forced Attention for Behavior Analysis
Tanay Agrawal, Michal Balazia, Philipp Müller +1
Human behavior understanding requires looking at minute details in the large context of a scene containing multiple input modalities. It is necessary as it allows the design of mor…