activity
20242026
most citedCounterfactual experience augmented off-policy reinforcement learning

5 citations · 5 across the 10 of their papers we have counts for

collaborators
Showing 2025Show all

7 papers · 1 filter

cs.CL2025

Probing the Difficulty Perception Mechanism of Large Language Models

Sunbowen Lee, Qingyu Yin, Chak Tou Leong +5

Large language models (LLMs) are increasingly deployed on complex reasoning tasks, yet little is known about their ability to internally evaluate problem difficulty, which is an es…

cs.SD2025

PoolingVQ: A VQVAE Variant for Reducing Audio Redundancy and Boosting Multi-Modal Fusion in Music Emotion Analysis

Dinghao Zou, Yicheng Gong, Xiaokang Li +2

Multimodal music emotion analysis leverages both audio and MIDI modalities to enhance performance. While mainstream approaches focus on complex feature extraction networks, we prop…

cs.SD2025

Emoanti: audio anti-deepfake with refined emotion-guided representations

Xiaokang Li, Yicheng Gong, Dinghao Zou +2

Audio deepfake is so sophisticated that the lack of effective detection methods is fatal. While most detection systems primarily rely on low-level acoustic features or pretrained s…

cs.CL2025

M-MRE: Extending the Mutual Reinforcement Effect to Multimodal Information Extraction

Chengguang Gan, Zhixi Cai, Yanbin Wei +3

Mutual Reinforcement Effect (MRE) is an emerging subfield at the intersection of information extraction and model interpretability. MRE aims to leverage the mutual understanding be…

cs.LG2025★ 5 cited

Counterfactual experience augmented off-policy reinforcement learning

Sunbowen Lee, Yicheng Gong, Chao Deng

Reinforcement learning control algorithms face significant challenges due to out-of-distribution and inefficient exploration problems. While model-based reinforcement learning enha…

cs.CL2025

Quantification of Large Language Model Distillation

Sunbowen Lee, Junting Zhou, Chang Ao +11

Model distillation is a fundamental technique in building large language models (LLMs), transferring knowledge from a teacher model to a student model. However, distillation can le…