1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents
Deeraj S K, Sadhana Devarajan, Krishna Mehra +1
Reinforcement learning from verifiable emotion rewards RLVER has produced language models with strong empathetic performance, evaluated on benchmarks that assume cooperative, hones…
cs.AI2024★ 1 cited
STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack
Shashank Kirtania, Naman Gupta, Priyanshu Gupta +7
Large Language Models (LLMs) often generate incorrect or outdated information, especially in low-resource settings or when dealing with private data. To address this, Retrieval-Aug…