2 papers
cs.SE2025
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
Florian Tambon, Amin Nikanjam, Cyrine Zid +2
Large Language Models (LLMs) excel in code-related tasks like code generation, but benchmark evaluations often overlook task characteristics, such as difficulty. Moreover, benchmar…
cs.SE2025
Performance Smells in ML and Non-ML Python Projects: A Comparative Study
François Belias, Leuson Da Silva, Foutse Khomh +1
Python is widely adopted across various domains, especially in Machine Learning (ML) and traditional software projects. Despite its versatility, Python is susceptible to performanc…