8 citations · 13 across the 14 of their papers we have counts for
Showing 2025 · cs.AIShow all
2 papers · 2 filters
cs.AI2025
To See or To Read: User Behavior Reasoning in Multimodal LLMs
Tianning Dong, Luyi Ma, Varun Vasudevan +3
Multimodal Large Language Models (MLLMs) are reshaping how modern agentic systems reason over sequential user-behavior data. However, whether textual or image representations of us…
cs.AI2025
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
Tao Zhang, Kehui Yao, Luyi Ma +7
Evaluating large language models (LLMs) as judges is increasingly critical for building scalable and trustworthy evaluation pipelines. We present ScalingEval, a large-scale benchma…