2 papers
cs.CL2026
When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning
Ruotao Xu, Yixin Ji, Yu Luo +5
Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of the underlying language mod…
cs.SE2026
DeepGuard: Secure Code Generation via Multi-Layer Semantic Aggregation
Li Huang, Zhongxin Liu, Yifan Wu +6
Large Language Models (LLMs) for code generation can replicate insecure patterns from their training data. To mitigate this, a common strategy for security hardening is to fine-tun…