4 papers
PROPHET: An Inferable Future Forecasting Benchmark with Causal Intervened Likelihood Estimation
Zhengwei Tao, Pu Wu, Zhi Jin +8
Predicting future events based on news on the Web stands as one of the ultimate aspirations of artificial intelligence. Recent advances in large language model (LLM)-based systems…
CodeMEM: AST-Guided Adaptive Memory for Repository-Level Iterative Code Generation
Peiding Wang, Li Zhang, Fang Liu +2
Large language models (LLMs) substantially enhance developer productivity in repository-level code generation through interactive collaboration. However, as interactions progress,…
GraphCodeAgent: Dual Graph-Guided LLM Agent for Retrieval-Augmented Repo-Level Code Generation
Jia Li, Xianjie Shi, Kechi Zhang +10
Writing code requires significant time and effort in software development. To automate this process, researchers have made substantial progress for code generation. Recently, large…
LONGCODEU: Benchmarking Long-Context Language Models on Long Code Understanding
Jia Li, Xuyuan Guo, Lei Li +7
Current advanced long-context language models offer great potential for real-world software engineering applications. However, progress in this critical domain remains hampered by…