5 papers
AdaStop: Cost-Aware Early Stopping for DNN Test Selection
Bonan Shen, Wei-Jung Huang, Xin Liu +2
Existing methods for testing deep neural networks (DNNs) primarily prioritize test inputs likely to reveal model faults under a fixed labeling budget. In practice, choosing that bu…
LLM-Driven CI-CD Workflow Intelligence for Cyber Systems Engineering
Bonan Shen, Jiazhou Gao, Tao Ning +2
CI/CD workflows have become executable operational policy: they decide what gets built, tested, released, and deployed, and they mediate how maintainers interact with delivery infr…
Context-Masked Truncated Reasoning Audits for Answer-Key Dependence in LLM Tutors
Bonan Shen, Dingyan Shang, Youting Wang +2
Large language model (LLM) tutors may have access to teacher notes, answer keys, rubrics, or retrieved solutions while producing student-facing explanations. We study whether trunc…
SemHash-LLM: A Multi-Granularity Semantic Hashing Framework for Document Deduplication
Xinyi Fang, Kejian Tong, Jiabei Liu +2
Large scale document deduplication must preserve semantic equivalence while remaining efficient over massive corpora. We present SemHash LLM, a multi granularity framework that uni…
Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking
Bonan Shen, Youting Wang, Dingyan Shang +1
Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut while the written reasoning st…