67 citations · 67 across the 2 of their papers we have counts for
9 papers
EgoForce: Forearm-Guided Camera-Space 3D Hand Pose from a Monocular Egocentric Camera
Christen Millerdurai, Shaoxiang Wang, Yaxu Xie +3
Reconstructing the absolute 3D pose and shape of the hands from the user's viewpoint using a single head-mounted camera is crucial for practical egocentric interaction in AR/VR, te…
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
Danish Nazir, Marcus Liwicki, Didier Stricker +1
Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guidance images to recover depth at…
TinyIceNet: Low-Power SAR Sea Ice Segmentation for On-Board FPGA Inference
Mhd Rashed Al Koutayni, Mohamed Selim, Gerd Reis +2
Accurate sea ice mapping is essential for safe maritime navigation in polar regions, where rapidly changing ice conditions require timely and reliable information. While Sentinel-1…
Inpaint360GS: Efficient Object-Aware 3D Inpainting via Gaussian Splatting for 360° Scenes
Shaoxiang Wang, Shihong Zhang, Christen Millerdurai +3
Despite recent advances in single-object front-facing inpainting using NeRF and 3D Gaussian Splatting (3DGS), inpainting in complex 360° scenes remains largely underexplored. This…
Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation
Chun-Peng Chang, Chen-Yu Wang, Julian Schmidt +2
Recent advancements in video generation have substantially improved visual quality and temporal coherence, making these models increasingly appealing for applications such as auton…
3D Spatial Understanding in MLLMs: Disambiguation and Evaluation
Chun-Peng Chang, Alain Pagani, Didier Stricker
Multimodal Large Language Models (MLLMs) have made significant progress in tasks such as image captioning and question answering. However, while these models can generate realistic…