From the 1 of 11 linked papers with an AI index.
11 papers
Detecting LLM-Generated Tokens in Human--LLM Coauthored Text
Yangjun Lu, Hongyi Zhou, Fabian Spill +3
The rise of human-AI collaborative writing has created a growing need for fine-grained detection methods that support localizing likely LLM-generated content in mixed-authorship do…
Segmenting Human-LLM Co-authored Text via Change Point Detection
Mengchu Li, Jin Zhu, Jinglai Li +1
The paper introduces algorithms that locate human-written versus LLM-generated segments within a mixed text by treating the problem as a change‑point detection task, and provides t…
Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning
Shijin Gong, Kai Ye, Jin Zhu +3
Recent advances in large language models (LLMs) have increasingly relied on reinforcement learning (RL) to improve their reasoning capabilities. Three types of approaches have been…
Perturbation is All You Need for Extrapolating Language Models
Zetai Cen, Jin Zhu, Xinwei Shen +1
This paper develops a statistical theory of extrapolation for large language models, by reinterpreting them through pre-post-additive noise models. In contrast to the standard auto…
Demystifying Group Relative Policy Optimization: Its Policy Gradient is a U-Statistic
Hongyi Zhou, Kai Ye, Erhan Xu +4
Group relative policy optimization (GRPO), a core methodological component of DeepSeekMath and DeepSeek-R1, has emerged as a cornerstone for scaling reasoning capabilities of large…
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
Hongyi Zhou, Jin Zhu, Kai Ye +3
Modern large language models (LLMs) such as GPT, Claude, and Gemini have transformed the way we learn, work, and communicate. Yet, their ability to produce highly human-like text r…