2 papers
cs.CL2026
PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations
Yuhe Wu, Guangyu Wang, Yuran Chen +6
As large language models (LLMs) evolve from conversational assistants into agents capable of handling complex tasks, they are increasingly deployed in high-risk domains. However, e…
cs.CE2026
BizCompass: Benchmarking the Reasoning Capabilities of LLMs in Business Knowledge and Applications
Jianing Hao, Yuhe Wu, Yuanjian Xu +5
Large language models (LLMs) hold great promise for business applications, yet business analysis remains inherently complex, demanding rigorous reasoning and the integration of div…