1 paper · 1 filter
Yiming Huang, Jianwen Luo, Yan Yu +8
We introduce DA-Code, a code generation benchmark specifically designed to assess LLMs on agent-based data science tasks. This benchmark features three core elements: First, the ta…