3 papers
cs.SE2026
Adaptive Hierarchical Evaluation of LLMs and SAST tools for CWE Prediction in Python
Muntasir Adnan, Carlos C. N. Kuhn
Large Language Models have become integral to software development, yet they frequently generate vulnerable code. Existing code vulnerability detection benchmarks employ binary cla…
cs.SE2025
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
Muntasir Adnan, Carlos C. N. Kuhn
The effectiveness of AI debugging follows a predictable exponential decay pattern; most models lose 60-80% of their debugging capability within just 2-3 attempts, despite iterative…
cs.SE2025
Large Language Model Guided Self-Debugging Code Generation
Muntasir Adnan, Zhiwei Xu, Carlos C. N. Kuhn
Automated code generation is gaining significant importance in intelligent computer programming and system deployment. However, current approaches often face challenges in computat…