Showing cs.SEShow all
2 papers · 1 filter
cs.SE2025
SWE-Compass: Towards Unified Evaluation of Agentic Coding Abilities for Large Language Models
Jingxuan Xu, Ken Deng, Weihao Li +36
Evaluating large language models (LLMs) for software engineering has been limited by narrow task coverage, language bias, and insufficient alignment with real-world developer workf…
cs.SE2025
SK2Decompile: LLM-based Two-Phase Binary Decompilation from Skeleton to Skin
Hanzhuo Tan, Weihao Li, Xiaolong Tian +4
Large Language Models (LLMs) have emerged as a promising approach for binary decompilation. However, the existing LLM-based decompilers still are somewhat limited in effectively pr…