1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2025★ 1 cited
Evaluation Framework for AI Systems in "the Wild"
Sarah Jabbour, Trenton Chang, Anindya Das Antar +13
Generative AI (GenAI) models have become vital across industries, yet current evaluation methods have not adapted to their widespread use. Traditional evaluations often rely on ben…
cs.AI2025
UFO2: The Desktop AgentOS
Chaoyun Zhang, He Huang, Chiming Ni +18
Recent Computer-Using Agents (CUAs), powered by multimodal large language models (LLMs), offer a promising direction for automating complex desktop workflows through natural langua…
q-bio.TO2023
The diagnostic utility of endocytoscopy for the detection of esophageal lesions: a systematic review and meta-analysis
Lu Wang, Bofu Tang, Feifei Liu +2
Objective: To systematically evaluate the value of endocytoscopy (ECS) in the diagnosis of early esophageal cancer (EC). Methods: Pubmed, Ovid and EMbase databases were searched to…