10 papers
Representation Robustness Under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving
Sagnik Nath, Edith Aurora Graf, Liang Zhang +1
Large language models (LLMs) are increasingly evaluated on mathematical problem solving, yet prior work often treats representationally equivalent formulations as interchangeable a…
BEMEval-Doc2Schema: Benchmarking Large Language Models for Structured Data Extraction in Building Energy Modeling
Yiyuan Jia, Xiaoqin Fu, Liang Zhang
Recent advances in foundation models, including large language models (LLMs), have created new opportunities to automate building energy modeling (BEM). However, systematic evaluat…
A State-Transition Framework for Efficient LLM Reasoning
Liang Zhang, Yu Zhao, Longyue Wang +4
While Long Chain-of-Thought (CoT) reasoning significantly improves Large Language Models (LLMs) performance on complex reasoning tasks, the substantial computational and memory cos…
Improved LLM Agents for Financial Document Question Answering
Nelvin Tan, Zian Seng, Liang Zhang +3
Large language models (LLMs) have shown impressive capabilities on numerous natural language processing tasks. However, LLMs still struggle with numerical question answering for fi…
Ask, Answer, and Detect: Role-Playing LLMs for Personality Detection with Question-Conditioned Mixture-of-Experts
Yifan Lyu, Liang Zhang
Understanding human personality is crucial for web applications such as personalized recommendation and mental health assessment. Existing studies on personality detection predomin…
Unilaw-R1: A Large Language Model for Legal Reasoning with Reinforcement Learning and Iterative Inference
Hua Cai, Shuang Zhao, Liang Zhang +5
Reasoning-focused large language models (LLMs) are rapidly evolving across various domains, yet their capabilities in handling complex legal problems remains underexplored. In this…