collaborators

8 papers

cs.CL2026

SDARE-Bench: Evaluating Large Language Models on Conversational Stigma Detection and Response in Dyadic and Group Dialogue

Stephanie Fong, Yiwen Jiang, Zimu Wang +12

Large Language Models (LLMs) are increasingly used in advice seeking and decision making that may affect social judgements. Despite stigma's profound effects on people and communit…

cs.AI2026

VIBE-Bench: Evaluating Personalized Large Language Models When Profiles Don't Mean Preferences

Yiwen Jiang, Yang Deng, Stephanie Fong +9

Personalized Large Language Models (PLLMs) aim to tailor responses to individual users, where a central challenge is preference reasoning: inferring query-relevant preferences from…

cs.CL2026

AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-Risk Symptom Measurement

Guilherme C. Oliveira, Stephanie Fong, Zimu Wang +10

Progress on AI for psychosis-risk assessment is limited by a data-access bottleneck. Real clinical interviews are difficult to share because of privacy, governance, and consent con…

cs.CV2026

NeRD: Neuro-Symbolic Rule Distillation for Efficient Ontology-Grounded Chain-of-Thought in Medical Image Diagnosis

Hongxi Yang, Yiwen Jiang, Siyuan Yan +8

Interpretability is essential for trustworthy medical image diagnosis. However, existing concept-driven interpretable methods have key limitations: Concept Bottleneck Models (CBMs)…

cs.SD2026

AudioProcessBench: Benchmark for Identifying Process Errors in Audio-Grounded Reasoning

Xiangyu Zhao, Junyu Yan, Yaling Shen +7

Large audio-language models (LALMs) increasingly use explicit reasoning traces for complex audio understanding, yet the evaluation of reasoning quality remains underexplored. Altho…

cs.CL2026

Do No Harm: Exposing Hidden Vulnerabilities of LLMs via Persona-based Client Simulation Attack in Psychological Counseling

Qingyang Xu, Yaling Shen, Stephanie Fong +7

The increasing use of large language models (LLMs) in mental healthcare raises safety concerns in high-stakes therapeutic interactions. A key challenge is distinguishing therapeuti…