collaborators

12 papers

cs.LG2026

Interpretable GOHR Agents via Sparse Autoencoders

Shiwei Tan, Yusong Zhao, Weiyi Qin +6

A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help explain their behavior. We rep…

cs.LG2026

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization

Hao Wang, Kun Yuan, Wenlin Zhong +4

Open-weight language models from different families exhibit complementary capabilities, motivating their consolidation into a compact student through on-policy distillation (OPD).…

cs.AI2026

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

Hao Wang, Jindong Han, Wei Fan +1

Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of analytical tools have created sign…

cs.AI2026

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

Dezhi Kong, Zhengzhao Feng, Qiliang Liang +12

Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confined to a reactive paradigm, where…

cs.CL2026

Compressing Sequences in the Latent Embedding Space: -Token Merging for Large Language Models

Zihao Xu, John Harvill, Ziwei Fan +3

Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically with input length. Token com…

cs.CL2026

Incentivizing Parametric Knowledge via Reinforcement Learning with Verifiable Rewards for Cross-Cultural Entity Translation

Jiang Zhou, Xiaohu Zhao, Xinwei Wu +8

Cross-cultural entity translation remains challenging for large language models (LLMs) as literal or phonetic renderings are usually yielded instead of culturally appropriate trans…