collaborators

6 papers

cs.AI2026

HardcoreLogic: Challenging Large Reasoning Models with Long-tail Logic Puzzle Games

Jingcong Liang, Shijun Wan, Xuehai Wu +5

Large Reasoning Models (LRMs) have demonstrated impressive performance on complex tasks, including logical puzzle games that require deriving solutions satisfying all constraints.…

cs.LG2026

Not All Models Suit Expert Offloading: On Local Routing Consistency of Mixture-of-Expert Models

Jingcong Liang, Siyuan Wang, Miren Tian +3

Mixture-of-Experts (MoE) enables efficient scaling of large language models (LLMs) with sparsely activated experts during inference. To effectively deploy large MoE models on memor…

cs.CL2024

From Individual to Society: A Survey on Social Simulation Driven by Large Language Model-based Agents

Xinyi Mou, Xuanwen Ding, Qi He +8

Traditional sociological research often relies on human participation, which, though effective, is expensive, challenging to scale, and with ethical concerns. Recent advancements i…

cs.CL2024

AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios

Xinyi Mou, Jingcong Liang, Jiayu Lin +8

Large language models (LLMs) are increasingly leveraged to empower autonomous agents to simulate human beings in various fields of behavioral research. However, evaluating their ca…

cs.CL2024

Overview of the CAIL 2023 Argument Mining Track

Jingcong Liang, Junlong Wang, Xinyu Zhai +16

We give a detailed overview of the CAIL 2023 Argument Mining Track, one of the Chinese AI and Law Challenge (CAIL) 2023 tracks. The main goal of the track is to identify and extrac…

cs.CL2024

Debatrix: Multi-dimensional Debate Judge with Iterative Chronological Analysis Based on LLM

Jingcong Liang, Rong Ye, Meng Han +4

How can we construct an automated debate judge to evaluate an extensive, vibrant, multi-turn debate? This task is challenging, as judging a debate involves grappling with lengthy t…