2 papers
cs.AI2026
Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability
Jeonghwan Choi, Taewon Yun, Minjeong Ban +3
Retrieval-augmented generation improves the factuality of large language models by grounding responses in retrieved evidence, yet existing evaluation frameworks struggle to provide…
cs.AI2026
What Makes a Sale? Simulating End-to-End Seller--Buyer Retail Dynamics with LLM Agents
Jeonghwan Choi, Jibin Hwang, Gyeonghun Sun +4
Evaluating retail strategies before deployment is difficult, as outcomes are determined across multiple stages, from seller-side persuasion through buyer-seller interaction to purc…