collaborators

7 papers

cs.CV2025

CCAD: Compressed Global Feature Conditioned Anomaly Detection

Xiao Jin, Liang Diao, Qixin Xiao +4

Anomaly detection holds considerable industrial significance, especially in scenarios with limited anomalous data. Currently, reconstruction-based and unsupervised representation-b…

cs.CV2025

Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy Optimization

Boyang Liu, Yifan Hu, Senjie Jin +5

Multimodal large language models (MLLMs) are well suited to image aesthetic assessment, as they can capture high-level aesthetic features leveraging their cross-modal understanding…

cs.CL2025

Entropy-Driven Pre-Tokenization for Byte-Pair Encoding

Yifan Hu, Frank Liang, Dachuan Zhao +4

Byte-Pair Encoding (BPE) has become a widely adopted subword tokenization method in modern language models due to its simplicity and strong empirical performance across downstream…

stat.ML2025

Reinforcement Learning with Continuous Actions Under Unmeasured Confounding

Yuhan Li, Eugene Han, Yifan Hu +4

This paper addresses the challenge of offline policy learning in reinforcement learning with continuous action spaces when unmeasured confounders are present. While most existing r…

cs.LG2025

Global Group Fairness in Federated Learning via Function Tracking

Yves Rychener, Daniel Kuhn, Yifan Hu

We investigate group fairness regularizers in federated learning, aiming to train a globally fair model in a distributed setting. Ensuring global fairness in distributed training p…

cs.CL2025

MPO: An Efficient Post-Processing Framework for Mixing Diverse Preference Alignment

Tianze Wang, Dongnan Gui, Yifan Hu +2

Reinforcement Learning from Human Feedback (RLHF) has shown promise in aligning large language models (LLMs). Yet its reliance on a singular reward model often overlooks the divers…