2 papers
cs.CL2026
DeepQuestion: Systematic Generation of Real-World Challenges for Evaluating LLMs Performance
Ali Khoramfar, Ali Ramezani, Mohammad Mahdi Mohajeri +3
While Large Language Models (LLMs) achieve near-human performance on standard benchmarks, their capabilities often fail to generalize to complex, real-world problems. To bridge thi…
cs.AI2025
A Dual-Axis Taxonomy of Knowledge Editing for LLMs: From Mechanisms to Functions
Amir Mohammad Salehoof, Ali Ramezani, Yadollah Yaghoobzadeh +1
Large language models (LLMs) acquire vast knowledge from large text corpora, but this information can become outdated or inaccurate. Since retraining is computationally expensive,…