collaborators

5 papers

cs.CL2026

Cross-Task Dissociation in Frontier Vision-Language Model Theory of Mind

Kejia Zhang, Youran Sun, Chugang Yi +1

Do frontier vision-language models present a coherent Theory-of-Mind (ToM) profile across tasks, matching the same human reference group, or does that profile fragment from one par…

cs.CL2026

Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?

Hai Hu, Siyuan Song, Chongtian Shao +3

In this paper, we push the boundary of LLM reasoning by testing them in a Chinese language game, xiehouyu, with novel xiehouyu created by linguists that had not existed before to a…

cs.CL2026

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting

Youran Sun, Xingyu Ren, Kejia Zhang +2

Real-world LLM applications are moving beyond single-agent workflows toward orchestrated multi-agent systems, yet current models still struggle to determine what each sub-agent nee…

cs.SE2026

Agon: An Autonomous Large-Scale Omnidisciplinary Research System Built on Prompt Economy

Youran Sun, Xingyu Ren, Chugang Yi +4

Large language models are making research production scalable, shifting the bottleneck from producing artifacts to judging claims. We present \textsc{Agon}, a research orchestrator…

cs.CL2025

A Systematic Assessment of Language Models with Linguistic Minimal Pairs in Chinese

Yikang Liu, Yeting Shen, Hongao Zhu +9

We present ZhoBLiMP, the largest linguistic minimal pair benchmark for Chinese, with over 100 paradigms, ranging from topicalization to the \textit{Ba} construction. We then train…