works on

From the 1 of 18 linked papers with an AI index.

activity
20242026
most citedMetaSyn: A Benchmark for LLM Agents on Meta-Analysis Articles from Nature Portfolio

1 citations · 1 across the 10 of their papers we have counts for

collaborators
Showing cs.CLShow all

11 papers · 1 filter

cs.CL2026

Mitigating Identity Essentialism in LLM Agents with Longitudinal Life Trajectories

Hexi Wang, Yujia Zhou, Bangde Du +7

Large language models (LLMs) offer a scalable approach to social simulation, but their credibility depends on how agents are constructed. Existing methods can partially reproduce p…

cs.CL20261 cited

MetaSyn: A Benchmark for LLM Agents on Meta-Analysis Articles from Nature Portfolio

Anzhe Xie, Weihang Su, Yujia Zhou +3

Systematic review and meta-analysis is an important method for scientific research. It comprehensively studies target research questions by combining evidence from multiple indepen…

cs.CL2026

Ontology Memory-Augmented ASR Correction for Long Text-Speech Interleaved Conversations

Xinxin Li, Huiyao Chen, Meishan Zhang +6

Automatic speech recognition (ASR) correction has traditionally focused on isolated utterances or short local contexts. However, as text and speech become increasingly interleaved…

cs.CL2026

Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Models

Changyue Wang, Weihang Su, Qingyao Ai +5

Large language models (LLMs) are widely used to tackle complex tasks with autonomous workflows. Recently, reusable natural language skills have emerged as a popular paradigm to inj…

cs.CL2026

IS-CoT: Breaking the Long-form Generation Collapse via Interleaved Structural Thinking

Zechen Sun, Yuyang Sun, Zecheng Tang +6

Generating coherent and controllable long-form content remains a persistent challenge for Large Language Models (LLMs). While reasoning-enhanced models have demonstrated success in…

cs.CL2026

LexRubric: A Rubric-Guided Diagnostic Benchmark for Open-Ended Legal Tasks

Yifan Chen, Haitao Li, Yiran Hu +6

As large language models (LLMs) are increasingly applied to real-world legal tasks, evaluating the reliability of their open-ended legal responses has become essential. These tasks…