1 citations · 1 across the 11 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Owen-Shapley Policy Optimization: A Principled RL Algorithm for Generative Search LLMs
Abhijnan Nath, Alireza Bagheri Garakani, Tianchen Zhou +3
Large language models are increasingly trained via reinforcement learning for personalized recommendation tasks, but standard methods like GRPO rely on sparse, sequence-level rewar…
cs.AI2025
Learning "Partner-Aware" Collaborators in Multi-Party Collaboration
Abhijnan Nath, Nikhil Krishnaswamy
Large Language Models (LLMs) are increasingly being deployed in agentic settings where they act as collaborators with humans. Therefore, it is increasingly important to be able to…