2 papers
cs.SE2026
RepoGenesis: Benchmarking End-to-End Microservice Generation from Readme to Repository
Zhiyuan Peng, Xin Yin, Pu Zhao +7
Large language models and agents have achieved remarkable progress in code generation. However, existing benchmarks focus on isolated function/class-level generation (e.g., ClassEv…
cs.AI2025
YuLan-OneSim: Towards the Next Generation of Social Simulator with Large Language Models
Lei Wang, Heyang Gao, Xiaohe Bo +2
Leveraging large language model (LLM) based agents to simulate human social behaviors has recently gained significant attention. In this paper, we introduce a novel social simulato…