2 papers
cs.CV2026
ReLE: A Scalable System and Structured Benchmark for Diagnosing Capability Anisotropy in Chinese LLMs
Rui Fang, Jian Li, Wei Chen +4
Large Language Models (LLMs) have achieved rapid progress in Chinese language understanding, yet accurately evaluating their capabilities remains challenged by benchmark saturation…
cs.CL2025
Breaking Thought Patterns: A Multi-Dimensional Reasoning Framework for LLMs
Xintong Tang, Meiru Zhang, Shang Xiao +5
Large language models (LLMs) are often constrained by rigid reasoning processes, limiting their ability to generate creative and diverse responses. To address this, a novel framewo…