activity
20232026
most citedHI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

1 citations · 1 across the 3 of their papers we have counts for

collaborators

5 papers

cs.CL2026

The Language-Energy Divide: Measuring Energy Costs of Multilingual LLM Inference

Naihao Deng, Alissa Shen, Yiming Feng +5

Large language models (LLMs) are increasingly deployed in multilingual settings, yet the energy costs of serving these models across different languages remain poorly understood. W…

cs.CL2026

The Wrong Kind of Right: Quantifying and Localizing Misfired Alignment in LLMs

Naihao Deng, Yiming Feng, Chimaobi Okite +4

Warning: This paper studies stereotypes and biases, and contains potentially disturbing examples, used for illustration purposes only. Our findings should not be interpreted as an…

cs.CL2025

QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model

Zongxian Yang, Jiayu Qian, Kay Chen Tan +4

Large language models (LLMs) face significant challenges in specialized biomedical tasks due to the inherent complexity of medical reasoning and the sensitive nature of clinical da…

cs.LG2024

Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs

Naihao Deng, Zhenjie Sun, Ruiqi He +5

In this paper, we investigate the effectiveness of various LLMs in interpreting tabular data through different prompting strategies and data formats. Our analyses extend across six…

cs.CL20231 cited

HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Yinghui He, Yufan Wu, Yilin Jia +3

Theory of Mind (ToM) is the ability to reason about one's own and others' mental states. ToM plays a critical role in the development of intelligence, language understanding, and c…