activity
20232026
collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV2025

TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering

Ayan Banerjee, Josep Llados, Umapada Pal +1

Text-to-story visualization is challenging due to the need for consistent interaction among multiple characters across frames. Existing methods struggle with character consistency,…

cs.CV2025

Privacy-enhancing Sclera Segmentation Benchmarking Competition: SSBC 2025

Matej Vitek, Darian Tomašević, Abhijit Das +32

This paper presents a summary of the 2025 Sclera Segmentation Benchmarking Competition (SSBC), which focused on the development of privacy-preserving sclera-segmentation models tra…

cs.CV2025

Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation

Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2

Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action within the scene. While the task is…

cs.CV2024

Decorrelation-based Self-Supervised Visual Representation Learning for Writer Identification

Arkadip Maitra, Shree Mitra, Siladittya Manna +2

Self-supervised learning has developed rapidly over the last decade and has been applied in many areas of computer vision. Decorrelation-based self-supervised pretraining has shown…

cs.CV2024

Correlation Weighted Prototype-based Self-Supervised One-Shot Segmentation of Medical Images

Siladittya Manna, Saumik Bhattacharya, Umapada Pal

Medical image segmentation is one of the domains where sufficient annotated data is not available. This necessitates the application of low-data frameworks like few-shot learning.…

cs.CV2024

FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting

Alloy Das, Sanket Biswas, Umapada Pal +2

The proliferation of scene text in both structured and unstructured environments presents significant challenges in optical character recognition (OCR), necessitating more efficien…