3 papers
cs.LG2026
The Law of Multi-Model Collaboration: Scaling Limits of Model Ensembling for Large Language Models
Dakuan Lu, Jiaqi Zhang, Cheng Yuan +2
Recent advances in large language models (LLMs) have been largely driven by scaling laws for individual models, which predict performance improvements as model parameters and data…
cs.AI2026
GUI-Eyes: Tool-Augmented Perception for Visual Grounding in GUI Agents
Chen Chen, Jiawei Shao, Dakuan Lu +4
Recent advances in vision-language models (VLMs) and reinforcement learning (RL) have driven progress in GUI automation. However, most existing methods rely on static, one-shot vis…
cs.AI2026
ScRPO: From Errors to Insights
Lianrui Li, Dakuan Lu, Jiawei Shao +1
We introduce Self-correction Relative Policy Optimization (ScRPO), a novel reinforcement learning framework designed to empower large language models with advanced mathematical rea…