3 papers
cs.CL2025
Time Awareness in Large Language Models: Benchmarking Fact Recall Across Time
David Herel, Vojtech Bartek, Jiri Jirak +1
Who is the US President? The answer changes depending on when the question is asked. While large language models (LLMs) are evaluated on various reasoning tasks, they often miss a…
cs.CL2025
Large Language Models: A Survey
Shervin Minaee, Tomas Mikolov, Narjes Nikzad +4
Large Language Models (LLMs) have drawn a lot of attention due to their strong performance on a wide range of natural language tasks, since the release of ChatGPT in November 2022.…
cs.CL2024
Thinking Tokens for Language Modeling
David Herel, Tomas Mikolov
How much is 56 times 37? Language models often make mistakes in these types of difficult calculations. This is usually explained by their inability to perform complex reasoning. Si…