3 citations · 5 across the 12 of their papers we have counts for
Showing 2023Show all
3 papers · 1 filter
cs.CL2023★ 3 cited
MM-BigBench: Evaluating Multimodal Models on Multimodal Content Comprehension Tasks
Xiaocui Yang, Wenfang Wu, Shi Feng +7
The popularity of multimodal large language models (MLLMs) has triggered a recent surge in research efforts dedicated to evaluating these models. Nevertheless, existing evaluation…
cs.AI2023
T-COL: Generating Counterfactual Explanations for General User Preferences on Variable Machine Learning Systems
Ming Wang, Daling Wang, Wenfang Wu +2
To address the interpretability challenge in machine learning (ML) systems, counterfactual explanations (CEs) have emerged as a promising solution. CEs are unique as they provide w…
cs.CL2023
RoCar: A Relationship Network-based Evaluation Method for Large Language Models
Ming Wang, Wenfang Wu, Chongyun Gao +3
Large language models (LLMs) have received increasing attention. However, due to the complexity of its capabilities, how to rationally evaluate the capabilities of LLMs is still a…