4 papers
Are MLMs Trapped in the Visual Room?
Yazhou Zhang, Chunwang Zou, Qimeng Liu +6
Can multi-modal large models (MLMs) that can ``see'' an image be said to ``understand'' it? Drawing inspiration from Searle's Chinese Room, we propose the \textbf{Visual Room} argu…
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
Ben Yao, Qiuchi Li, Yazhou Zhang +4
While LLMs have demonstrated medical knowledge and conversational ability, their deployment in clinical practice raises new risks: patients may place greater trust in LLM-generated…
Beyond Single-Sentence Prompts: Upgrading Value Alignment Benchmarks with Dialogues and Stories
Yazhou Zhang, Qimeng Liu, Qiuchi Li +2
Evaluating the value alignment of large language models (LLMs) has traditionally relied on single-sentence adversarial prompts, which directly probe models with ethically sensitive…
Commander-GPT: Fully Unleashing the Sarcasm Detection Capability of Multi-Modal Large Language Models
Yazhou Zhang, Chunwang Zou, Bo Wang +1
Sarcasm detection, as a crucial research direction in the field of Natural Language Processing (NLP), has attracted widespread attention. Traditional sarcasm detection tasks have t…