1 paper
Aabid Karim, Abdul Karim, Bhoomika Lohana +3
We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google, Meta, DeepSeek, Mistral, and Mic…