2 papers
cs.SE2025
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
Florian Tambon, Amin Nikanjam, Cyrine Zid +2
Large Language Models (LLMs) excel in code-related tasks like code generation, but benchmark evaluations often overlook task characteristics, such as difficulty. Moreover, benchmar…
cs.SE2025
One Documentation Does Not Fit All: Case Study of TensorFlow Documentation
Sharuka Promodya Thirimanne, Elim Yoseph Lemango, Giulio Antoniol +1
Software documentation guides the proper use of tools or services. With the rapid growth of machine learning libraries, individuals from various fields are incorporating machine le…