3 papers
cs.LG2026
A Tutorial on Diffusion Theory: From Differential Equations to Diffusion Models
Jiayi Fu, Yuxia Wang
Diffusion models have emerged as a dominant framework for generative modeling, but their mathematical foundations are often presented separately through diffusion probabilistic mod…
cs.CR2026
Can LLM Infer Risk Information From MCP Server System Logs?
Jiayi Fu, Yuansen Zhang, Yinggui Wang
Large Language Models (LLMs) demonstrate strong capabilities in solving complex tasks when integrated with external tools. The Model Context Protocol (MCP) has become a standard in…
cs.LG2026
Reward Shaping to Mitigate Reward Hacking in RLHF
Jiayi Fu, Xuandong Zhao, Chengyuan Yao +3
Reinforcement learning from human feedback (RLHF) is widely used to align large language models (LLMs) with human preferences. However, RLHF remains vulnerable to \emph{reward hack…