1 paper
Andrea Gaggioli, Giuseppe Casaburi, Leonardo Ercolani +3
This study investigates the reliability and validity of five advanced Large Language Models (LLMs), Claude 3.5, DeepSeek v2, Gemini 2.5, GPT-4, and Mistral 24B, for automated essay…