2 papers
cs.CL2026
Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective
Yukun Chen, Xinyu Zhang, Boyi Deng +6
As large language models (LLMs) are employed worldwide, existing evaluation paradigms for their multilingual capabilities primarily focus on factual task performance, neglecting th…
cs.CV2026
UrbanAlign: Post-hoc Semantic Calibration for VLM-Human Preference Alignment
Yecheng Zhang, Rong Zhao, Zhizhou Sha +10
Vision-language models (VLMs) can describe urban scenes in rich detail, yet consistently fail to produce reliable human preference labels in domain-specific tasks such as safety as…