Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Closed-Loop Neural Activation Control in Vision-Language-Action Models
Abhijith Babu, Ramneet Kaur, Nathaniel D. Bastian +5
Vision-Language-Action (VLA) models can be steered at test time by intervening on semantically meaningful internal directions, but existing methods use a fixed steering coefficient…
cs.AI2024
Towards a Game-theoretic Understanding of Explanation-based Membership Inference Attacks
Kavita Kumari, Murtuza Jadliwala, Sumit Kumar Jha +1
Model explanations improve the transparency of black-box machine learning (ML) models and their decisions; however, they can also be exploited to carry out privacy threats such as…