10 papers · 1 filter
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
Ayan Banerjee, Josep Llados, Umapada Pal +1
Text-to-story visualization is challenging due to the need for consistent interaction among multiple characters across frames. Existing methods struggle with character consistency,…
Privacy-enhancing Sclera Segmentation Benchmarking Competition: SSBC 2025
Matej Vitek, Darian Tomašević, Abhijit Das +32
This paper presents a summary of the 2025 Sclera Segmentation Benchmarking Competition (SSBC), which focused on the development of privacy-preserving sclera-segmentation models tra…
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2
Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action within the scene. While the task is…
Decorrelation-based Self-Supervised Visual Representation Learning for Writer Identification
Arkadip Maitra, Shree Mitra, Siladittya Manna +2
Self-supervised learning has developed rapidly over the last decade and has been applied in many areas of computer vision. Decorrelation-based self-supervised pretraining has shown…
Correlation Weighted Prototype-based Self-Supervised One-Shot Segmentation of Medical Images
Siladittya Manna, Saumik Bhattacharya, Umapada Pal
Medical image segmentation is one of the domains where sufficient annotated data is not available. This necessitates the application of low-data frameworks like few-shot learning.…
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
Alloy Das, Sanket Biswas, Umapada Pal +2
The proliferation of scene text in both structured and unstructured environments presents significant challenges in optical character recognition (OCR), necessitating more efficien…