3 papers
cs.CV2026
Beyond Static Anchors: Bounded Prototype Conditioning for Language-Free Medical Anomaly Detection
Yibo Wan, Jinyu Cai, See-kiong Ng
Medical anomaly detection identifies abnormal images and localizes lesions under scarce supervision while generalizing across organs and modalities. Existing CLIP-based methods red…
cs.CL2025
More Than One Teacher: Adaptive Multi-Guidance Policy Optimization for Diverse Exploration
Xiaoyang Yuan, Yujuan Ding, Yi Bin +5
Reinforcement Learning with Verifiable Rewards (RLVR) is a promising paradigm for enhancing the reasoning ability in Large Language Models (LLMs). However, prevailing methods prima…
cs.AI2025
WereWolf-Plus: An Update of Werewolf Game setting Based on DSGBench
Xinyuan Xia, Yuanyi Song, Haomin Ma +1
With the rapid development of LLM-based agents, increasing attention has been given to their social interaction and strategic reasoning capabilities. However, existing Werewolf-bas…