3 papers
cs.CL2026
Investigating LLM Capabilities on Long Context Comprehension for Medical Question Answering
Feras AlMannaa, Talia Tseriotou, Jenny Chim +1
This study is the first to investigate LLM comprehension capabilities over long-context (LC), clinically relevant medical Question Answering (QA) beyond MCQA. Our comprehensive app…
cs.SE2025
BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
Terry Yue Zhuo, Minh Chien Vu, Jenny Chim +30
Task automation has been greatly empowered by the recent advances in Large Language Models (LLMs) via Python code, where the tasks ranging from software engineering development to…
cs.CY2024
Large language models for mental health
Andreas Triantafyllopoulos, Yannik Terhorst, Iosif Tsangko +10
Digital technologies have long been explored as a complement to standard procedure in mental health research and practice, ranging from the management of electronic health records…