4 papers
Beyond Static Anchors: Bounded Prototype Conditioning for Language-Free Medical Anomaly Detection
Yibo Wan, Jinyu Cai, See-kiong Ng +1
Medical anomaly detection identifies abnormal images and localizes lesions under scarce supervision while generalizing across organs and modalities. Existing CLIP-based methods red…
GLOW: Graph-Language Co-Reasoning for Agentic Workflow Performance Prediction
Wei Guan, Jian Cao, Jinyu Cai +3
Agentic Workflows (AWs) have emerged as a promising paradigm for solving complex tasks. However, the scalability of automating their generation is severely constrained by the high…
More Than One Teacher: Adaptive Multi-Guidance Policy Optimization for Diverse Exploration
Xiaoyang Yuan, Yujuan Ding, Yi Bin +5
Reinforcement Learning with Verifiable Rewards (RLVR) is a promising paradigm for enhancing the reasoning ability in Large Language Models (LLMs). However, prevailing methods prima…
WereWolf-Plus: An Update of Werewolf Game setting Based on DSGBench
Xinyuan Xia, Yuanyi Song, Haomin Ma +1
With the rapid development of LLM-based agents, increasing attention has been given to their social interaction and strategic reasoning capabilities. However, existing Werewolf-bas…