2 papers
cs.CV2025
An End-to-End Depth-Based Pipeline for Selfie Image Rectification
Ahmed Alhawwary, Janne Mustaniemi, Phong Nguyen-Ha +1
Portraits or selfie images taken from a close distance typically suffer from perspective distortion. In this paper, we propose an end-to-end deep learning-based rectification pipel…
cs.CV2025
Towards an Automated Multimodal Approach for Video Summarization: Building a Bridge Between Text, Audio and Facial Cue-Based Summarization
Md Moinul Islam, Sofoklis Kakouros, Janne Heikkilä +1
The increasing volume of video content in educational, professional, and social domains necessitates effective summarization techniques that go beyond traditional unimodal approach…