7 papers
Trust Regions Sell, But Who's Buying? Overlap Geometry as an Alternative Trust Region for Policy Optimization
Gaurish Trivedi, Alakh Sharma, Kartikey Singh Bhandari +4
Standard trust-region methods constrain policy updates via Kullback-Leibler (KL) divergence. However, KL controls only an average divergence and does not directly prevent rare, lar…
Beyond Single Bugs: Benchmarking Large Language Models for Multi-Vulnerability Detection
Chinmay Pushkar, Sanchit Kabra, Dhruv Kumar +1
Large Language Models (LLMs) have demonstrated significant potential in automated software security, particularly in vulnerability detection. However, existing benchmarks primarily…
Talking with Oompa Loompas: A novel framework for evaluating linguistic acquisition of LLM agents
Sankalp Tattwadarshi Swain, Anshika Krishnatray, Dhruv Kumar +1
Existing evaluation studies on linguistic competence of large language models (LLM agents) have focused primarily on vocabulary learning, morphological rule induction, syntactic ge…
HAEPO: History-Aggregated Exploratory Policy Optimization
Gaurish Trivedi, Alakh Sharma, Kartikey Singh Bhandari +3
Exploration is essential in modern learning, from reinforcement learning environments with small neural policies to large language models (LLMs). Existing work, such as DPO, levera…
SAC: A Framework for Measuring and Inducing Personality Traits in LLMs with Dynamic Intensity Control
Adithya Chittem, Aishna Shrivastava, Sai Tarun Pendela +2
Large language models (LLMs) have gained significant traction across a wide range of fields in recent years. There is also a growing expectation for them to display human-like pers…
The Impact of Large Language Models on K-12 Education in Rural India: A Thematic Analysis of Student Volunteer's Perspectives
Harshita Goyal, Garima Garg, Prisha Mordia +3
AI-driven education, particularly Large Language Models (LLMs), has the potential to address learning disparities in rural K-12 schools. However, research on AI adoption in rural I…