3 papers
cs.LG2026
Discovering Reinforcement Learning Interfaces with Large Language Models
Akshat Singh Jaswal, Ashish Baghel, Paras Chopra
Reinforcement learning systems rely on environment interfaces that specify observations and reward functions, yet constructing these interfaces for new tasks often requires substan…
cs.CR2026
AWE: Adaptive Agents for Dynamic Web Penetration Testing
Akshat Singh Jaswal, Ashish Baghel
Modern web applications are increasingly produced through AI-assisted development and rapid no-code deployment pipelines, widening the gap between accelerating software velocity an…
cs.CL2025
It Takes Two: A Dual Stage Approach for Terminology-Aware Translation
Akshat Singh Jaswal
This paper introduces DuTerm, a novel two-stage architecture for terminology-constrained machine translation. Our system combines a terminology-aware NMT model, adapted via fine-tu…