activity
20242026
collaborators

10 papers

cs.CV2026

Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation

Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2

Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action within the scene. While the task is…

cs.GR2025

d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining

Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2

Structural guidance in an image-to-image translation allows intricate control over the shapes of synthesized images. Generating high-quality realistic images from user-specified ro…

cs.CV2025

A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing

Prasun Roy, Subhankar Ghosh, Umapada Pal

Air-writing refers to virtually writing linguistic characters through hand gestures in three-dimensional space with six degrees of freedom. This paper proposes a generic video came…

cs.CV2025

Semantically Consistent Person Image Generation

Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2

We propose a data-driven approach for context-aware person image generation. Specifically, we attempt to generate a person image such that the synthesized instance can blend into a…

cs.CV2025

TIPS: Text-Induced Pose Synthesis

Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya +2

In computer vision, human pose synthesis and transfer deal with probabilistic image generation of a person in a previously unseen pose from an already available observation of that…

cs.CV2025

Scene Aware Person Image Generation through Global Contextual Conditioning

Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya +2

Person image generation is an intriguing yet challenging problem. However, this task becomes even more difficult under constrained situations. In this work, we propose a novel pipe…