4 papers
CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing
Mingzhi Zhu, Michele Merler, Raju Pavuluri +1
Code agents must both reason over long-horizon repository state and obey strict tool-use protocols. In paired Instruct/Thinking checkpoints, these capabilities are complementary bu…
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
Advait Pavuluri, Bridget McGinn, Ashita Saxena +6
Java remains central to enterprise software, and many applications outlive their original architecture. Migrating them across frameworks is a behavior-preserving refactoring spanni…
Multi-task Code LLMs: Data Mix or Model Merge?
Mingzhi Zhu, Boris Sobolev, Rahul Krishna +3
Recent research advocates deploying smaller, specialized code LLMs in agentic frameworks alongside frontier models, sparking interest in efficient strategies for multi-task learnin…
Usage, Effects and Requirements for AI Coding Assistants in the Enterprise: An Empirical Study
Maja Vukovic, Rangeet Pan, Tin Kam Ho +3
The rise of large language models (LLMs) has accelerated the development of automated techniques and tools for supporting various software engineering tasks, e.g., program understa…