Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Amplification Effects in Test-Time Reinforcement Learning: Safety and Reasoning Vulnerabilities
Vanshaj Khattar, Md Rafi ur Rashid, Moumita Choudhury +4
Test-time training (TTT) has recently emerged as a promising method to improve the reasoning abilities of large language models (LLMs), in which the model directly learns from test…
cs.LG2025
Detecting Zero-Day Attacks in Digital Substations via In-Context Learning
Faizan Manzoor, Vanshaj Khattar, Akila Herath +5
The occurrences of cyber attacks on the power grids have been increasing every year, with novel attack techniques emerging every year. In this paper, we address the critical challe…
cs.LG2024
Optimization Solution Functions as Deterministic Policies for Offline Reinforcement Learning
Vanshaj Khattar, Ming Jin
Offline reinforcement learning (RL) is a promising approach for many control applications but faces challenges such as limited data coverage and value function overestimation. In t…