2 papers
cs.LG2026
Are Latent Reasoning Models Easily Interpretable?
Connor Dilgren, Sarah Wiegreffe
Latent reasoning models (LRMs) have attracted significant research interest due to their low inference cost (relative to explicit reasoning models) and theoretical ability to explo…
cs.CR2026
SecRepoBench: Benchmarking Code Agents for Secure Code Completion in Real-World Repositories
Chihao Shen, Connor Dilgren, Purva Chiniya +3
This paper introduces SecRepoBench, a benchmark to evaluate code agents on secure code completion in real-world repositories. SecRepoBench has 318 code completion tasks in 27 C/C++…