1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.SE2025
CUARewardBench: A Benchmark for Evaluating Reward Models on Computer-using Agent
Haojia Lin, Xiaoyu Tan, Yulei Qin +9
Computer-using agents (CUAs) enable task completion through natural interaction with operating systems and software interfaces. While script-based verifiers are widely adopted for…
cs.AI2025★ 1 cited
FlowAgent: Achieving Compliance and Flexibility for Workflow Agents
Yuchen Shi, Siqi Cai, Zihan Xu +7
The integration of workflows with large language models (LLMs) enables LLM-based agents to execute predefined procedures, enhancing automation in real-world applications. Tradition…
cs.CV2024
Leveraging Open Knowledge for Advancing Task Expertise in Large Language Models
Yuncheng Yang, Yulei Qin, Tong Wu +9
The cultivation of expertise for large language models (LLMs) to solve tasks of specific areas often requires special-purpose tuning with calibrated behaviors on the expected stabl…