2 papers
cs.CR2025
Jailbreaking in the Haystack
Rishi Rajesh Shah, Chen Henry Wu, Shashwat Saxena +3
Recent advances in long-context language models (LMs) have enabled million-token inputs, expanding their capabilities across complex tasks like computer-use agents. Yet, the safety…
cs.RO2025
SITCOM: Scaling Inference-Time COMpute for VLAs
Ayudh Saxena, Harsh Shah, Sandeep Routray +2
Learning robust robotic control policies remains a major challenge due to the high cost of collecting labeled data, limited generalization to unseen environments, and difficulties…