3 papers
cs.AI2026
X-RAY: Mapping LLM Reasoning Capability via Formalized and Calibrated Probes
Tianxi Gao, Yufan Cai, Yusi Yuan +1
Large language models (LLMs) achieve promising performance, yet their ability to reason remains poorly understood. Existing evaluations largely emphasize task-level accuracy, often…
cs.SE2026
Generalizing Test Cases for Comprehensive Test Scenario Coverage
Binhang Qi, Yun Lin, Xinyi Weng +4
Test cases are essential for software development and maintenance. In practice, developers derive multiple test cases from an implicit pattern based on their understanding of requi…
cs.LG2025
Supervised Robustness-preserving Data-free Neural Network Pruning
Mark Huasong Meng, Guangdong Bai, Sin Gee Teo +1
When deploying pre-trained neural network models in real-world applications, model consumers often encounter resource-constraint platforms such as mobile and smart devices. They ty…