From the 1 of 5 linked papers with an AI index.
5 papers
ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning
Debjyoti Das Adhikary, Aritra Hazra, Partha Pratim Chakrabarti
Improving video captioning quality typically demands retraining large vision-language models, an expensive and often impractical requirement. Existing training-free alternatives in…
BackgroundMellow: A Multi-Modal Cohesive Framework for Narrative-Driven Rich Cinematic Soundscape Generation
Ajitesh Jamulkar, Aritra Hazra
The paper introduces BackgroundMellow, a multi‑modal framework that converts long‑form textual narratives into cohesive, cinematic soundscapes by decomposing text into audio cues,…
Direction-Conditioned Policies via Compositional Subgoal Scoring for Online Goal-Conditioned Reinforcement Learning
Swaminathan S K, Damiya Gondha, Theyanesh Eswaramoorthy Rajahkrishnan +1
Hamilton-Jacobi-Bellman theory implies that the optimal goal-conditioned action depends on the goal only through the gradient of the goal-reaching distance at the current state, ye…
SPAARS: Safer RL Policy Alignment through Abstract Exploration and Refined Exploitation of Action Space
Swaminathan S K, Aritra Hazra
Offline-to-online reinforcement learning (RL) offers a promising paradigm for robotics by pre-training policies on safe, offline demonstrations and fine-tuning them via online inte…
ReFrame: Rectification Framework for Image Explaining Architectures
Debjyoti Das Adhikary, Aritra Hazra, Partha Pratim Chakrabarti
Image explanation has been one of the key research interests in the Deep Learning field. Throughout the years, several approaches have been adopted to explain an input image fed by…