2 citations · 6 across the 6 of their papers we have counts for
6 papers
GUI Testing Arena: A Unified Benchmark for Advancing Autonomous GUI Testing Agent
Kangjia Zhao, Jiahui Song, Leigang Sha +5
Nowadays, research on GUI agents is a hot topic in the AI community. However, current research focuses on GUI task automation, limiting the scope of applications in various GUI sce…
CLEVA: Chinese Language Models EVAluation Platform
Yanyang Li, Jianqiao Zhao, Duo Zheng +8
With the continuous emergence of Chinese Large Language Models (LLMs), how to evaluate a model's capabilities has become an increasingly significant issue. The absence of a compreh…
Provable Multi-instance Deep AUC Maximization with Stochastic Pooling
Dixian Zhu, Bokun Wang, Zhi Chen +4
This paper considers a novel application of deep AUC maximization (DAM) for multi-instance learning (MIL), in which a single class label is assigned to a bag of instances (e.g., mu…
On the Structural Generalization in Text-to-SQL
Jieyu Li, Lu Chen, Ruisheng Cao +5
Exploring the generalization of a text-to-SQL parser is essential for a system to automatically adapt the real-world databases. Previous works provided investigations focusing on l…
Hyper-relationship Learning Network for Scene Graph Generation
Yibing Zhan, Zhi Chen, Jun Yu +3
Generating informative scene graphs from images requires integrating and reasoning from various graph components, i.e., objects and relationships. However, current scene graph gene…
Few-Shot NLU with Vector Projection Distance and Abstract Triangular CRF
Su Zhu, Lu Chen, Ruisheng Cao +3
Data sparsity problem is a key challenge of Natural Language Understanding (NLU), especially for a new target domain. By training an NLU model in source domains and applying the mo…