18 citations · 18 across the 11 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
BEAVER: An Efficient Deterministic LLM Verifier
Tarun Suresh, Nalin Wadhwa, Debangshu Banerjee +1
As large language models (LLMs) transition from research prototypes to production systems, practitioners often need reliable methods to verify model outputs and characterize tail r…
cs.AI2024
Towards Reliable Alignment: Uncertainty-aware RLHF
Debangshu Banerjee, Aditya Gopalan
Recent advances in aligning Large Language Models with human preferences have benefited from larger reward models and better preference data. However, most of these methodologies r…
cs.AI2022
Critic Algorithms using Cooperative Networks
Debangshu Banerjee, Kavita Wagh
An algorithm is proposed for policy evaluation in Markov Decision Processes which gives good empirical results with respect to convergence rates. The algorithm tracks the Projected…