1 citations · 1 across the 3 of their papers we have counts for
3 papers · 1 filter
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation
Ahmed Abul Hasanaath, Bicheng Xu, Mir Rayat Imtiaz Hossain +2
Gloss-free Sign Language Translation (SLT) translates sign language videos into spoken-language sentences without gloss annotations, avoiding costly labeling but requiring fine-gra…
USTM: Unified Spatial and Temporal Modeling for Continuous Sign Language Recognition
Ahmed Abul Hasanaath, Hamzah Luqman
Continuous sign language recognition (CSLR) requires precise spatio-temporal modeling to accurately recognize sequences of gestures in videos. Existing frameworks often rely on CNN…
FSBI: Deepfakes Detection with Frequency Enhanced Self-Blended Images
Ahmed Abul Hasanaath, Hamzah Luqman, Raed Katib +1
Advances in deepfake research have led to the creation of almost perfect manipulations undetectable by human eyes and some deepfakes detection tools. Recently, several techniques h…