2 papers
cs.CL2026
As Easy as Rocket Science: Assessing the Ability of Large Language Models to Interpret Negation in Figurative Language
Jasmine Owers, Edwin Simpson, Martha Lewis
Figurative language and negation are two areas that challenge current language models, however, both are widely used throughout written and spoken language. Large language models (…
cs.CL2024
Evaluating the Robustness of Analogical Reasoning in Large Language Models
Martha Lewis, Melanie Mitchell
LLMs have performed well on several reasoning benchmarks, including ones that test analogical reasoning abilities. However, there is debate on the extent to which they are performi…