10 papers
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2
Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action within the scene. While the task is…
d-Sketch: Improving Visual Fidelity of Sketch-to-Image Translation with Pretrained Latent Diffusion Models without Retraining
Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2
Structural guidance in an image-to-image translation allows intricate control over the shapes of synthesized images. Generating high-quality realistic images from user-specified ro…
A CNN Based Framework for Unistroke Numeral Recognition in Air-Writing
Prasun Roy, Subhankar Ghosh, Umapada Pal
Air-writing refers to virtually writing linguistic characters through hand gestures in three-dimensional space with six degrees of freedom. This paper proposes a generic video came…
Semantically Consistent Person Image Generation
Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh +2
We propose a data-driven approach for context-aware person image generation. Specifically, we attempt to generate a person image such that the synthesized instance can blend into a…
TIPS: Text-Induced Pose Synthesis
Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya +2
In computer vision, human pose synthesis and transfer deal with probabilistic image generation of a person in a previously unseen pose from an already available observation of that…
Scene Aware Person Image Generation through Global Contextual Conditioning
Prasun Roy, Subhankar Ghosh, Saumik Bhattacharya +2
Person image generation is an intriguing yet challenging problem. However, this task becomes even more difficult under constrained situations. In this work, we propose a novel pipe…