works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

cs.CV2026

ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning

Debjyoti Das Adhikary, Aritra Hazra, Partha Pratim Chakrabarti

Improving video captioning quality typically demands retraining large vision-language models, an expensive and often impractical requirement. Existing training-free alternatives in…

cs.LG2026

BackgroundMellow: A Multi-Modal Cohesive Framework for Narrative-Driven Rich Cinematic Soundscape Generation

Ajitesh Jamulkar, Aritra Hazra

The paper introduces BackgroundMellow, a multi‑modal framework that converts long‑form textual narratives into cohesive, cinematic soundscapes by decomposing text into audio cues,…

cs.LG2026

Direction-Conditioned Policies via Compositional Subgoal Scoring for Online Goal-Conditioned Reinforcement Learning

Swaminathan S K, Damiya Gondha, Theyanesh Eswaramoorthy Rajahkrishnan +1

Hamilton-Jacobi-Bellman theory implies that the optimal goal-conditioned action depends on the goal only through the gradient of the goal-reaching distance at the current state, ye…

cs.LG2026

SPAARS: Safer RL Policy Alignment through Abstract Exploration and Refined Exploitation of Action Space

Swaminathan S K, Aritra Hazra

Offline-to-online reinforcement learning (RL) offers a promising paradigm for robotics by pre-training policies on safe, offline demonstrations and fine-tuning them via online inte…

cs.CV2025

ReFrame: Rectification Framework for Image Explaining Architectures

Debjyoti Das Adhikary, Aritra Hazra, Partha Pratim Chakrabarti

Image explanation has been one of the key research interests in the Deep Learning field. Throughout the years, several approaches have been adopted to explain an input image fed by…