2 papers
cs.CL2026
MIS-Bench: Benchmarking Multimodal LLMs for Psychotherapeutic Interpersonal Skills Assessment
Yuhan Lu, Yi Yao, Hua Shen +2
Multimodal large language models (MLLMs) are increasingly used as evaluators, yet their reliability in professional assessment tasks that require expert judgment remains unclear. W…
cs.AI2026
ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms
Alexandria K. Vail, Marcelo Cicconet, Katie Aafjes-van Doorn +2
Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic Decomposition for Automated Pr…