2 papers
cs.CL2025
CoreEval: Automatically Building Contamination-Resilient Datasets with Real-World Knowledge toward Reliable LLM Evaluation
Jingqian Zhao, Bingbing Wang, Geng Tu +5
Data contamination poses a significant challenge to the fairness of LLM evaluations in natural language processing tasks by inadvertently exposing models to test data during traini…
cs.MM2025
Reply with Sticker: New Dataset and Model for Sticker Retrieval
Bin Liang, Bingbing Wang, Zhixin Bai +6
Using stickers in online chatting is very prevalent on social media platforms, where the stickers used in the conversation can express someone's intention/emotion/attitude in a viv…