Publications (9)
How Do VLMs Fail? Vision-Operation Misalignment in Compositional VQA
Navya Gupta, Bingjie Xu, Avinash Anand +2
Compositional visual question answering requires Vision-Language Models (VLMs) to execute multiple reasoning operations like object selection, spatial relation resolution, and attr…
Variable Rate Lossy Source-Channel Coding over Channels with Feedback
Timothy Liu, Fady Alajaji, Tamás Linder
A variable-rate lossy joint source-channel coding scheme for burst-noise communication channels with noiseless feedback is introduced. The scheme comprises a multi-stage channel op…
PACE: Persona Adaptation through Conversational Elicitation in Human-Robot Interaction
Peizhen Li, Longbing Cao, Megani Rajendran +3
Equipping humanoid robots with coherent and adaptable personas is crucial for fostering natural, engaging, and trustworthy human-robot interaction (HRI). However, existing approach…
IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning
Navya Gupta, Rishitej Reddy Vyalla, Avinash Anand +8
Curriculum learning helps language models tackle complex reasoning by gradually increasing task difficulty. However, it often fails to generate consistent step-by-step reasoning, e…
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes
Avinash Anand, Mahisha Ramesh, Avni Mittal +8
Reasoning has become central to how Large Language Models (LLMs) are evaluated and interpreted, spanning Chain-of-Thought (CoT), mathematical problem-solving, multi-hop question an…
RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations
Parmitha Vangapandu, Sai Ganesh Mokkapati, Sathwik Narkedimilli +4
In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relationships to capture both menta…