2 papers
cs.CL2024
Chumor 2.0: Towards Benchmarking Chinese Humor Understanding
Ruiqi He, Yushu He, Longju Bai +7
Existing humor datasets and evaluations predominantly focus on English, leaving limited resources for culturally nuanced humor in non-English languages like Chinese. To address thi…
cs.LG2024
Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs
Naihao Deng, Zhenjie Sun, Ruiqi He +5
In this paper, we investigate the effectiveness of various LLMs in interpreting tabular data through different prompting strategies and data formats. Our analyses extend across six…