From the 1 of 5 linked papers with an AI index.
5 papers
CORE-Bench: A Comprehensive Benchmark for Code Retrieval in the Era of Agentic Coding
Fuwei Zhang, Yanzhao Zhang, Mingxin Li +5
The paper presents CORE-Bench, a large benchmark designed to evaluate code retrieval tasks needed by coding agents, covering code understanding, issue-to-edit localization, and bro…
SimGym: A Framework for A/B Test Simulation in E-Commerce with Traffic-Grounded VLM Agents
Han Li, Vibhor Malik, Zahra Zanjani Foumani +17
A/B testing remains the gold standard for evaluating modifications to e-commerce storefronts, yet it diverts traffic, requires weeks to reach statistical significance, and risks de…
SynGR: Unleashing the Potential of Cross-Modal Synergy for Generative Recommendation
Wei Chen, Xingyu Guo, Shuang Li +6
Generative Recommendation (GR) has emerged as a promising paradigm by formulating item recommendation as a sequence-to-sequence generation task over item identifiers. Recent studie…
Multi-Aspect Cross-modal Quantization for Generative Recommendation
Fuwei Zhang, Xiaoyu Liu, Dongbo Xi +5
Generative Recommendation (GR) has emerged as a new paradigm in recommender systems. This approach relies on quantized representations to discretize item features, modeling users'…
CAT-ID: Category-Tree Integrated Document Identifier Learning for Generative Retrieval In E-commerce
Xiaoyu Liu, Fuwei Zhang, Yiqing Wu +6
Generative retrieval (GR) has gained significant attention as an effective paradigm that integrates the capabilities of large language models (LLMs). It generally consists of two s…