8 papers
PiSAs: Benchmarking Contextual Integrity in Multi-User Agentic Systems
Shubham Gupta, Nazanin Mohammadi Sepahvand, Abhinav Kumar +6
As LLM agents evolve from single-user assistants into shared organizational infrastructure, new privacy risks emerge: inappropriate information may not only be exposed through outp…
Enhancing Audio Captioning with Auxiliary AudioSet Semantics
Shubham Gupta, Adarsh Arigala, Sri Rama Murty Kodukula
Automatic Audio Captioning (AAC) seeks to generate natural language descriptions of complex acoustic scenes, bridging auditory perception and language understanding. However, word-…
Shared Representation Learning for Reference-Guided Targeted Sound Detection
Shubham Gupta, Adarsh Arigala, B. R. Dilleswari +1
Human listeners exhibit the remarkable ability to segregate a desired sound from complex acoustic scenes through selective auditory attention, motivating the study of Targeted Soun…
Hierarchical Retrieval at Scale: Bridging Transparency and Efficiency
Shubham Gupta, Zichao Li, Tianyi Chen +4
Information retrieval is a core component of many intelligent systems as it enables conditioning of outputs on new and large-scale datasets. While effective, the standard practice…
Joint Multimodal Contrastive Learning for Robust Spoken Term Detection and Keyword Spotting
Ramesh Gundluru, Shubham Gupta, Sri Rama Murty K
Acoustic Word Embeddings (AWEs) improve the efficiency of speech retrieval tasks such as Spoken Term Detection (STD) and Keyword Spotting (KWS). However, existing approaches suffer…
Audio Prototypical Network For Controllable Music Recommendation
Fırat Ãncel, Emiliano Penaloza, Haolun Wu +4
Traditional recommendation systems represent user preferences in dense representations obtained through black-box encoder models. While these models often provide strong recommenda…