20 citations · 23 across the 6 of their papers we have counts for
3 papers · 1 filter
Measuring the Unmeasurable: Markov Chain Reliability for LLM Agents
Phat T. Tran-Truong, Xuan-Bach Le
Large language model (LLM) agents increasingly operate as sequential software systems, but their reliability is often summarized by scalar benchmark metrics. Metrics such as pass$@…
AutoPruner: Transformer-Based Call Graph Pruning
Thanh Le-Cong, Hong Jin Kang, Truong Giang Nguyen +4
Constructing a static call graph requires trade-offs between soundness and precision. Program analysis techniques for constructing call graphs are unfortunately usually imprecise.…
On Reliability of Patch Correctness Assessment
Xuan Bach D. Le, Lingfeng Bao, David Lo +2
Current state-of-the-art automatic software repair (ASR) techniques rely heavily on incomplete specifications, e.g., test suites, to generate repairs. This, however, may render ASR…