1 citations · 1 across the 12 of their papers we have counts for
4 papers · 1 filter
CPO: Evaluating Cross-Modal Composition and Counterfactual Performance in Omnimodal Models
Swapnanil Mukherjee, Agyeya Negi, Tanuja Ganu +1
Current Multimodal Large Language Models (MLLMs) can process diverse sensory inputs, yet their reasoning remains heavily biased toward a dominant modality, resulting in brittle cro…
I Can't Believe It's Corrupt: Evaluating Corruption in Multi-Agent Governance Systems
Vedanta S P, Ponnurangam Kumaraguru
Large language models are increasingly proposed as autonomous agents for high-stakes public workflows, yet we lack systematic evidence about whether they would follow institutional…
Small Models, Big Tasks: An Exploratory Empirical Study on Small Language Models for Function Calling
Ishan Kavathekar, Raghav Donakanti, Ponnurangam Kumaraguru +1
Function calling is a complex task with widespread applications in domains such as information retrieval, software engineering and automation. For example, a query to book the shor…
Wu's Method can Boost Symbolic AI to Rival Silver Medalists and AlphaGeometry to Outperform Gold Medalists at IMO Geometry
Shiven Sinha, Ameya Prabhu, Ponnurangam Kumaraguru +2
Proving geometric theorems constitutes a hallmark of visual reasoning combining both intuitive and logical skills. Therefore, automated theorem proving of Olympiad-level geometry p…