12 papers
Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning
Abhijith Babu, Ramneet Kaur, Vishal Pramanik +7
Multi-agent LLM systems can improve reasoning by pooling diverse perspectives, but their effectiveness depends on coordinating communication, particularly in hidden-profile setting…
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
Trilok Padhi, Ramneet Kaur, Krishiv Agarwal +9
Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, planning, and acting within interactive environments. Despite their growing capabi…
Bayesian Adaptation Gym: A Benchmark for the Bayesian Low-Rank Adaptation of Multi-Modal Language Models
Colin Samplawski, Ramneet Kaur, Manoj Acharya +2
Large multi-modal language models are increasingly deployed in high-stakes domains, making well-calibrated uncertainty essential. Traditional Bayesian methods approximate posterior…
Closed-Loop Neural Activation Control in Vision-Language-Action Models
Abhijith Babu, Ramneet Kaur, Nathaniel D. Bastian +5
Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use a fixed steering coefficient…
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
Krishiv Agarwal, Ramneet Kaur, Colin Samplawski +6
Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities rooted in model internals. We pr…
Do Diffusion Models Dream of Electric Planes? Discrete and Continuous Simulation-Based Inference for Aircraft Design
Aurelien Ghiglino, Daniel Elenius, Anirban Roy +7
In this paper, we generate conceptual engineering designs of electric vertical take-off and landing (eVTOL) aircraft. We follow the paradigm of simulation-based inference (SBI), wh…