4 papers
Episodic Memory Temporal Consistency for Cooperative Multi-Agent Reinforcement Learning
Zicheng Zhao, Yu Lan, Chengzhengxu Li +2
Cooperative Multi-Agent Reinforcement Learning (MARL) frequently suffers from severe reward sparsity and exploration bottlenecks. While episodic memory mechanisms mitigate these is…
DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection
Guoxin Ma, Xiaoming Liu, Hongyang Chen +6
Detecting machine-generated text has become a critical challenge amid the rapid advancement of LLMs, yet existing detectors degrade severely under domain shift. Through systematic…
Can Reasoning Path still be Effective as Input? Bridging Post-Reasoning to Chain-of-Thought Compression
Chengzhengxu Li, Xiaoming Liu, Zhaohan Zhang +5
Recent developments have enabled advanced reasoning in Large Language Models (LLMs) via long Chain-of-Thought (CoT), trading efficiency during inference for performance. Existing w…
MGT-Prism: Enhancing Domain Generalization for Machine-Generated Text Detection via Spectral Alignment
Shengchao Liu, Xiaoming Liu, Chengzhengxu Li +4
Large Language Models have shown growing ability to generate fluent and coherent texts that are highly similar to the writing style of humans. Current detectors for Machine-Generat…