2 papers
cs.CL2026
Evaluating Large Language Models on Computer Science University Exams in Data Structures
Edan Gabay, Yael Maoz, Jonathan Stahl +7
We present a comprehensive evaluation of Large Language Models (LLMs) on Computer Science (CS) Data Structure examination questions. Our work introduces a new benchmark dataset com…
cs.AI2025
Using multi-agent architecture to mitigate the risk of LLM hallucinations
Abd Elrahman Amer, Magdi Amer
Improving customer service quality and response time are critical factors for maintaining customer loyalty and increasing a company's market share. While adopting emerging technolo…