1 paper
Angel Yanguas-Gil, Matthew T. Dearing, Jeffrey W. Elam +5
In this work we introduce an open-ended question benchmark, ALDbench, to evaluate the performance of large language models (LLMs) in materials synthesis, and in particular in the f…