activity
20212025
most citedHierarchical Context Pruning: Optimizing Real-World Code Completion with Repository-Level Pretrained Code LLMs

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL2025

SWE-Flow: Synthesizing Software Engineering Data in a Test-Driven Manner

Lei Zhang, Jiaxi Yang, Min Yang +6

We introduce **SWE-Flow**, a novel data synthesis framework grounded in Test-Driven Development (TDD). Unlike existing software engineering data that rely on human-submitted issues…

cs.CL2024

ExecRepoBench: Multi-level Executable Code Completion Evaluation

Jian Yang, Jiajun Zhang, Jiaxi Yang +9

Code completion has become an essential tool for daily software development. Existing evaluation benchmarks often employ static methods that do not fully capture the dynamic nature…

cs.CL20241 cited

Hierarchical Context Pruning: Optimizing Real-World Code Completion with Repository-Level Pretrained Code LLMs

Lei Zhang, Yunshui Li, Jiaming Li +7

Some recently developed code large language models (Code LLMs) have been pre-trained on repository-level code data (Repo-Code LLMs), enabling these models to recognize repository s…

cs.DS2023

Multi-dimensional Data Quick Query for Blockchain-based Federated Learning

Jiaxi Yang, Sheng Cao, Peng xiangLi +2

Due to the drawbacks of Federated Learning (FL) such as vulnerability of a single central server, centralized federated learning is shifting to decentralized federated learning, a…

cs.CL2021

A Survey of Toxic Comment Classification Methods

Kehan Wang, Jiaxi Yang, Hongjun Wu

While in real life everyone behaves themselves at least to some extent, it is much more difficult to expect people to behave themselves on the internet, because there are few check…