papers

Publications (10)

cs.LG2026

SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation

Wen Wang, Jiahua Bao, Tu Yongsiqi +8

We aim to improve model performance in multi-reward reinforcement learning training process. Existing Group reward-Decoupled Normalization Policy Optimization (GDPO) has mitigated…

cs.LG2024

HGOE: Hybrid External and Internal Graph Outlier Exposure for Graph Out-of-Distribution Detection

Junwei He, Qianqian Xu, Yangbangyan Jiang +3

With the progressive advancements in deep graph learning, out-of-distribution (OOD) detection for graph data has emerged as a critical challenge. While the efficacy of auxiliary da…

cs.LG2023

ADA-GAD: Anomaly-Denoised Autoencoders for Graph Anomaly Detection

Junwei He, Qianqian Xu, Yangbangyan Jiang +2

Graph anomaly detection is crucial for identifying nodes that deviate from regular behavior within graphs, benefiting various domains such as fraud detection and social network. Al…

cs.LG2026

ChronoMedicalWorld: A Medical World Model for Learning Patient Trajectories from Longitudinal Care Data

Jiangyuan Wang, Xuyong Chen, Junwei He +3

Long-horizon clinical simulation -- predicting how a patient's physiology evolves over years under specified interventions -- is central to chronic-disease care, yet existing elect…

cs.SD2025

TAME: Temporal Audio-based Mamba for Enhanced Drone Trajectory Estimation and Classification

Zhenyuan Xiao, Huanran Hu, Guili Xu +1

The increasing prevalence of compact UAVs has introduced significant risks to public safety, while traditional drone detection systems are often bulky and costly. To address these…

cs.IT2024

Symbol Detection for Coarsely Quantized OTFS

Junwei He, Haochuan Zhang, Chao Dong +1

This paper explicitly models a coarse and noisy quantization in a communication system empowered by orthogonal time frequency space (OTFS) for cost and power efficiency. We first p…