activity
20242026
collaborators

13 papers

cs.RO2026

No Free Checker: A Survey of Verifiers for Robot Policies

Yang Wan, Xihang Yue, Zhirui Liu +7

A verifier for robot policies reads a candidate behavior and returns a score for how well it did, used both to evaluate vision-language-action policies and to train them. Verifiers…

cs.AI2026

From Interaction Traces to Persistent Skills: Online Evolution for Computer-Use Agents

Longtao Hu, Xiao Liang, Linchao Zhu

Computer-use agents can execute increasingly complex tasks in graphical interfaces, but their interaction experience is typically transient: procedural knowledge acquired from one…

cs.AI2026

SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents

Yang Wan, Zhenhao Zhang, Jierui Wang +1

Deciding whether a trajectory actually fulfills its instruction governs how we measure computer-use agents on long-horizon graphical-user-interface tasks and how we train them with…

cs.AI2026

VISTA: View-Consistent Self-Verified Training for GUI Grounding

Xinyu Qiu, Yunzhu Zhang, Heng Jia +3

When applying Group Relative Policy Optimization (GRPO) for GUI Grounding, rollouts are sampled from a single screenshot view; groups often become either all failures on difficult…

physics.chem-ph2026

A Fixed-Point Neural Operator for Size- and Functional-Transferable Hamiltonian Prediction

Yunhong Lou, Xihang Yue, Xinran Wei +2

Predicting the Kohn-Sham Hamiltonian with machine learning can accelerate density functional theory while retaining access to molecular orbitals, energy levels, and electronic-stru…

cs.CV2026

GPD: Guided Progressive Distillation for Fast and High-Quality Video Generation

Xiao Liang, Yunzhu Zhang, Linchao Zhu

Diffusion models have achieved remarkable success in video generation; however, the high computational cost of the denoising process remains a major bottleneck. Existing approaches…