46 citations · 62 across the 4 of their papers we have counts for
5 papers · 1 filter
HeadPosr: End-to-end Trainable Head Pose Estimation using Transformer Encoders
Naina Dhingra
In this paper, HeadPosr is proposed to predict the head poses using a single RGB image. \textit{HeadPosr} uses a novel architecture which includes a transformer encoder. In concret…
LwPosr: Lightweight Efficient Fine-Grained Head Pose Estimation
Naina Dhingra
This paper presents a lightweight network for head pose estimation (HPE) task. While previous approaches rely on convolutional neural networks, the proposed network \textit{LwPosr}…
Border-SegGCN: Improving Semantic Segmentation by Refining the Border Outline using Graph Convolutional Network
Naina Dhingra, George Chogovadze, Andreas Kunz
We present Border-SegGCN, a novel architecture to improve semantic segmentation by refining the border outline using graph convolutional networks (GCN). The semantic segmentation n…
BGT-Net: Bidirectional GRU Transformer Network for Scene Graph Generation
Naina Dhingra, Florian Ritter, Andreas Kunz
Scene graphs are nodes and edges consisting of objects and object-object relationships, respectively. Scene graph generation (SGG) aims to identify the objects and their relationsh…
Res3ATN -- Deep 3D Residual Attention Network for Hand Gesture Recognition in Videos
Naina Dhingra, Andreas Kunz
Hand gesture recognition is a strenuous task to solve in videos. In this paper, we use a 3D residual attention network which is trained end to end for hand gesture recognition. Bas…