3 papers
cs.LG2026
Stochastic Linear Bandits with Partially Observed Actions
Gautam Dasarathy, Vineet Gattani, Lalit Jain
The stochastic linear bandit, where actions are represented as vectors and rewards are linear, is a central paradigm for sequential decision making. We study a partially observed v…
cs.SE2026
Toward Automated Validation of Language Model Synthesized Test Cases using Semantic Entropy
Hamed Taherkhani, Jiho Shin, Muhammad Ammar Tahir +3
Modern Large Language Model (LLM)-based programming agents often rely on test execution feedback to refine their generated code. These tests are synthetically generated by LLMs. Ho…
cs.LG2024
Communication-Efficient Federated Learning over Wireless Channels via Gradient Sketching
Vineet Sunil Gattani, Junshan Zhang, Gautam Dasarathy
Large-scale federated learning (FL) over wireless multiple access channels (MACs) has emerged as a crucial learning paradigm with a wide range of applications. However, its widespr…