8 papers
Closed-Loop Neural Activation Control in Vision-Language-Action Models
Abhijith Babu, Ramneet Kaur, Nathaniel D. Bastian +5
Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use a fixed steering coefficient…
Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs
Vishal Pramanik, Maisha Maliha, Nathaniel D. Bastian +1
Attribution methods seek to explain language model predictions by quantifying the contribution of input tokens to generated outputs. However, most existing techniques are designed…
Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion
Vishal Pramanik, Maisha Maliha, Susmit Jha +1
Large language models remain vulnerable to jailbreak attacks -- inputs designed to bypass safety mechanisms and elicit harmful responses -- despite advances in alignment and instru…
Grammar-Forced Translation of Natural Language to Temporal Logic using LLMs
William English, Dominic Simon, Sumit Kumar Jha +1
Translating natural language (NL) into a formal language such as temporal logic (TL) is integral for human communication with robots and autonomous systems. State-of-the-art approa…
Verifiable Natural Language to Linear Temporal Logic Translation: A Benchmark Dataset and Evaluation Suite
William H English, Chase Walker, Dominic Simon +2
Empirical evaluation of state-of-the-art natural-language (NL) to temporal-logic (TL) translation systems reveals near-perfect performance on existing benchmarks. However, current…
Circuit Partitioning Using Large Language Models for Quantum Compilation and Simulations
Pranav Sinha, Sumit Kumar Jha, Sunny Raj
We are in the midst of the noisy intermediate-scale quantum (NISQ) era, where quantum computers are limited by noisy gates, some of which are more error-prone than others and can r…