6 papers
Evaluating VLMs on Multimodal Aristotelian Persuasion Tasks
Khondoker Ittehadul Islam
Vision Language Models (VLMs) have demonstrated exceptional performance across various tasks. However, they have not yet been thoroughly evaluated on more complex tasks. The Persua…
Reveal-Bangla: A Dataset for Cross-Lingual Multi-Step Reasoning Evaluation
Khondoker Ittehadul Islam, Gabriele Sarti
Language models have demonstrated remarkable performance on complex multi-step reasoning tasks. However, their evaluation has been predominantly confined to high-resource languages…
What Really Counts? Examining Step and Token Level Attribution in Multilingual CoT Reasoning
Jeremias Ferrao, Ezgi Basar, Khondoker Ittehadul Islam +1
This study investigates the attribution patterns underlying Chain-of-Thought (CoT) reasoning in multilingual LLMs. While prior works demonstrate the role of CoT prompting in improv…
Joint Effects of Argumentation Theory, Audio Modality and Data Enrichment on LLM-Based Fallacy Classification
Hongxu Zhou, Hylke Westerdijk, Khondoker Ittehadul Islam
This study investigates how context and emotional tone metadata influence large language model (LLM) reasoning and performance in fallacy classification tasks, particularly within…
Improving OCR for Historical Texts of Multiple Languages
Hylke Westerdijk, Ben Blankenborg, Khondoker Ittehadul Islam
This paper presents our methodology and findings from three tasks across Optical Character Recognition (OCR) and Document Layout Analysis using advanced deep learning techniques. F…
Leveraging Sentiment for Offensive Text Classification
Khondoker Ittehadul Islam
In this paper, we conduct experiment to analyze whether models can classify offensive texts better with the help of sentiment. We conduct this experiment on the SemEval 2019 task 6…