Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation
Wenbo Zhang, Hengrui Cai, Wenyu Chen
Large language models (LLMs) have demonstrated significant utility in real-world applications, exhibiting impressive capabilities in natural language processing and understanding.…
cs.CL2024
Recognizing Limits: Investigating Infeasibility in Large Language Models
Wenbo Zhang, Zihang Xu, Hengrui Cai
Large language models (LLMs) have shown remarkable performance in various tasks but often fail to handle queries that exceed their knowledge and capabilities, leading to incorrect…