2 papers
cs.CL2026
Evidence-Bounded Mental Health Reasoning from Heterogeneous Speech Protocols
Chengyuan Gao, Jiang Wu, Tao Lu +4
Computational mental health screening using multimodal speech and text has shown great promise. However, existing models often assume all clinical speech protocols carry equivalent…
cs.AI2026
QuoteBench: How Matched Scores Can Hide Command-Path Failures
Shangao Li, Yao Zhang, Volker Tresp +1
LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot distinguish command-generation er…