activity
20242026
collaborators

20 papers

cs.LG2026

Iterative Feature Space Optimization through Incremental Adaptive Evaluation

Yanping Wu, Yanyong Huang, Zhengzhang Chen +5

Iterative feature space optimization involves systematically evaluating and adjusting the feature space to improve downstream task performance. However, existing works suffer from…

cs.LG2026

PageLLM: A Multi-Grained Reward Framework for Whole-Page Optimization with Large Language Models

Xinyuan Wang, Liang Wu, Dongjie Wang +1

Whole-page optimization (WPO) decides how search and recommendation results are surfaced to users, and large language models (LLMs) open a new route to it by treating page generati…

cs.CL2026

Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach

Hongyu Cao, Kunpeng Liu, Dongjie Wang +1

Large language models exhibit strong reasoning capabilities, yet often rely on shortcuts such as surface pattern matching and answer memorization rather than genuine logical infere…

cs.AI2026

AgentOS: From Application Silos to a Natural Language-Driven Data Ecosystem

Rui Liu, Tao Zhe, Dongjie Wang +5

The rapid emergence of open-source, locally hosted intelligent agents marks a critical inflection point in human-computer interaction. Systems such as OpenClaw demonstrate that Lar…

cs.LG2026

Sim2Act: Robust Simulation-to-Decision Learning via Adversarial Calibration and Group-Relative Perturbation

Hongyu Cao, Jinghan Zhang, Kunpeng Liu +5

Simulation-to-decision learning enables safe policy training in digital environments without risking real-world deployment, and has become essential in mission-critical domains suc…

cs.LG2026

Continuous Optimization for Feature Selection with Permutation-Invariant Embedding and Policy-Guided Search

Rui Liu, Rui Xie, Zijun Yao +2

Feature selection removes redundant features to enhanc performance and computational efficiency in downstream tasks. Existing works often struggle to capture complex feature intera…