Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
FullStack Bench: Evaluating LLMs as Full Stack Coders
Bytedance-Seed-Foundation-Code-Team, :, Yao Cheng +53
As the capabilities of code large language models (LLMs) continue to expand, their applications across diverse code intelligence domains are rapidly increasing. However, most exist…
cs.AI2024
Seed-CTS: Unleashing the Power of Tree Search for Superior Performance in Competitive Coding Tasks
Hao Wang, Boyi Liu, Yufeng Zhang +1
Competition-level code generation tasks pose significant challenges for current state-of-the-art large language models (LLMs). For example, on the LiveCodeBench-Hard dataset, model…