2 papers
cs.AI2026
Is Your LLM Really Mastering the Concept? A Multi-Agent Benchmark
Shuhang Xu, Weijian Deng, Yixuan Zhou +1
Concepts serve as fundamental abstractions that support human reasoning and categorization. However, it remains unclear whether large language models truly capture such conceptual…
cs.CL2025
CoMet: Metaphor-Driven Covert Communication for Multi-Agent Language Games
Shuhang Xu, Fangwei Zhong
Metaphors are a crucial way for humans to express complex or subtle ideas by comparing one concept to another, often from a different domain. However, many large language models (L…