3 papers
cs.SE2026
Human-aligned AI Model Cards with Weighted Hierarchy Architecture
Pengyue Yang, Haolin Jin, Qingwen Zeng +3
The proliferation of Large Language Models (LLMs) has led to a burgeoning ecosystem of specialized, domain-specific models. While this rapid growth accelerates innovation, it has s…
cs.CL2026
Trust in One Round: Confidence Estimation for Large Language Models via Structural Signals
Pengyue Yang, Jiawen Wen, Haolin Jin +3
Large language models (LLMs) are increasingly deployed in domains where errors carry high social, scientific, or safety costs. Yet standard confidence estimators, such as token lik…
cs.SE2025
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
Jiawen Wen, Bangshuo Zhu, Huaming Chen
Recent studies have demonstrated outstanding capabilities of large language models (LLMs) in software engineering tasks, including code generation and comprehension. While LLMs hav…