Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AgentWorld: Personality-Aware Reliability Evaluation for Agentic Information Retrieval
Gunja Agarwal, Arup Kumar Das, Arun Menon +2
Evaluation of agentic information retrieval remains limited to scripted interactions with uniform users, missing both natural personality diversity and adversarial brittleness. We…
cs.AI2026
State-Grounded Multi-Agent Synthetic Data Generation for Tool-Augmented LLMs
Rahul Khedar, Eshita, Sneha Teja Sree Reddy Thondapu +10
Training tool-augmented LLM agents requires large corpora of multi-turn, tool-grounded conversational data that is expensive to annotate, privacy-constrained in production settings…