3 papers
cs.CL2025
CoreEval: Automatically Building Contamination-Resilient Datasets with Real-World Knowledge toward Reliable LLM Evaluation
Jingqian Zhao, Bingbing Wang, Geng Tu +5
Data contamination poses a significant challenge to the fairness of LLM evaluations in natural language processing tasks by inadvertently exposing models to test data during traini…
cs.MM2024
Towards Real-World Stickers Use: A New Dataset for Multi-Tag Sticker Recognition
Bingbing Wang, Bin Liang, Chun-Mei Feng +6
In real-world conversations, the diversity and ambiguity of stickers often lead to varied interpretations based on the context, necessitating the requirement for comprehensively un…
cs.MM2024
Reply with Sticker: New Dataset and Model for Sticker Retrieval
Bin Liang, Bingbing Wang, Zhixin Bai +6
Using stickers in online chatting is very prevalent on social media platforms, where the stickers used in the conversation can express someone's intention/emotion/attitude in a viv…