Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents
Zeying Hao, Hao Guo, Mengtao Xu +5
Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult to articulate in text alone…
cs.AI2026
CMI-Mem: Toward Generalizable Long-Term Memory Management via CMI-Augmented Reinforcement Learning
Yubo Wang, Qiuyu Zhao, Zenghui Sun +6
Memory Manager models are pivotal in agent systems. Existing reinforcement-learning methods commonly use LLM-judged synthetic question-answer (QA) pairs: this provides useful downs…