13 papers
No Free Checker: A Survey of Verifiers for Robot Policies
Yang Wan, Xihang Yue, Zhirui Liu +7
A verifier for robot policies reads a candidate behavior and returns a score for how well it did, used both to evaluate vision-language-action policies and to train them. Verifiers…
From Interaction Traces to Persistent Skills: Online Evolution for Computer-Use Agents
Longtao Hu, Xiao Liang, Linchao Zhu
Computer-use agents can execute increasingly complex tasks in graphical interfaces, but their interaction experience is typically transient: procedural knowledge acquired from one…
SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents
Yang Wan, Zhenhao Zhang, Jierui Wang +1
Deciding whether a trajectory actually fulfills its instruction governs how we measure computer-use agents on long-horizon graphical-user-interface tasks and how we train them with…
VISTA: View-Consistent Self-Verified Training for GUI Grounding
Xinyu Qiu, Yunzhu Zhang, Heng Jia +3
When applying Group Relative Policy Optimization (GRPO) for GUI Grounding, rollouts are sampled from a single screenshot view; groups often become either all failures on difficult…
A Fixed-Point Neural Operator for Size- and Functional-Transferable Hamiltonian Prediction
Yunhong Lou, Xihang Yue, Xinran Wei +2
Predicting the Kohn-Sham Hamiltonian with machine learning can accelerate density functional theory while retaining access to molecular orbitals, energy levels, and electronic-stru…
GPD: Guided Progressive Distillation for Fast and High-Quality Video Generation
Xiao Liang, Yunzhu Zhang, Linchao Zhu
Diffusion models have achieved remarkable success in video generation; however, the high computational cost of the denoising process remains a major bottleneck. Existing approaches…