3 papers
cs.CL2025
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
Hessa A. Alawwad, Anas Zafar, Areej Alhothali +3
Multimodal large language models (MLLMs) have shown success in vision-language tasks, but their ability to reason over complex educational materials remains largely untested. This…
cs.IR2025
Beyond Retrieval: Joint Supervision and Multimodal Document Ranking for Textbook Question Answering
Hessa Alawwad, Usman Naseem, Areej Alhothali +2
Textbook question answering (TQA) is a complex task, requiring the interpretation of complex multimodal context. Although recent advances have improved overall performance, they of…
cs.CL2025
Enhancing textual textbook question answering with large language models and retrieval augmented generation
Hessa Abdulrahman Alawwad, Areej Alhothali, Usman Naseem +2
Textbook question answering (TQA) is a challenging task in artificial intelligence due to the complex nature of context needed to answer complex questions. Although previous resear…