Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
OS-MAP: How Far Can Computer-Using Agents Go in Breadth and Depth?
Xuetian Chen, Yinghao Chen, Xinfeng Yuan +12
Computer-using agents have shown strong potential to boost human productivity and enable new application forms across platforms. While recent advances have led to usable applicatio…
cs.AI2025
Evaluating the Correctness of Inference Patterns Used by LLMs for Judgment
Lu Chen, Yuxuan Huang, Yixing Li +6
This paper presents a method to analyze the inference patterns used by Large Language Models (LLMs) for judgment in a case study on legal LLMs, so as to identify potential incorrec…
cs.AI2025
Citrus: Leveraging Expert Cognitive Pathways in a Medical Language Model for Advanced Medical Decision Support
Guoxin Wang, Minyu Gao, Shuai Yang +9
Large language models (LLMs), particularly those with reasoning capabilities, have rapidly advanced in recent years, demonstrating significant potential across a wide range of appl…