3 papers
cs.AI2025
CrowdAgent: Multi-Agent Managed Multi-Source Annotation System
Maosheng Qin, Renyu Zhu, Mingxuan Xia +8
High-quality annotated data is a cornerstone of modern Natural Language Processing (NLP). While recent methods begin to leverage diverse annotation sources-including Large Language…
cs.LG2025
The Effects of Data Augmentation on Confidence Estimation for LLMs
Rui Wang, Renyu Zhu, Minmin Lin +4
Confidence estimation is crucial for reflecting the reliability of large language models (LLMs), particularly in the widely used closed-source models. Utilizing data augmentation f…
cs.LG2025
Digital Player: Evaluating Large Language Models based Human-like Agent in Games
Jiawei Wang, Kai Wang, Shaojie Lin +11
With the rapid advancement of Large Language Models (LLMs), LLM-based autonomous agents have shown the potential to function as digital employees, such as digital analysts, teacher…