3 papers
cs.AI2025
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
Tao Zhang, Kehui Yao, Luyi Ma +7
Evaluating large language models (LLMs) as judges is increasingly critical for building scalable and trustworthy evaluation pipelines. We present ScalingEval, a large-scale benchma…
cs.IR2025
CARTS: Collaborative Agents for Recommendation Textual Summarization
Jiao Chen, Kehui Yao, Reza Yousefi Maragheh +6
Current recommendation systems often require some form of textual data summarization, such as generating concise and coherent titles for product carousels or other grouped item dis…
cs.IR2024
Improving Sequential Recommender Systems with Online and In-store User Behavior
Luyi Ma, Aashika Padmanabhan, Anjana Ganesh +9
Online e-commerce platforms have been extending in-store shopping, which allows users to keep the canonical online browsing and checkout experience while exploring in-store shoppin…