3 papers
cs.LG2026
Fairness in two-player zero-sum games with bandit feedback
S Akash, Pratik Gajane
We study two-player zero-sum games (TPZSGs) with bandit feedback under fairness constraints requiring every action to be played with probability at least . Existing instance-d…
cs.IR2026
SUMMIR: A Hallucination-Aware Framework for Ranking Sports Insights from LLMs
Nitish Kumar, Sannu Kumar, S Akash +3
With the rapid proliferation of online sports journalism, extracting meaningful pre-game and post-game insights from articles is essential for enhancing user engagement and compreh…
cs.LG2026
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
S Akash, Pratik Gajane, Jawar Singh
Multi-dueling bandits, where a learner selects arms per round and observes only the winner, arise naturally in many applications including ranking and recommendation sys…