most citedRethinking Memory in LLM based Agents: Representations, Operations, and Emerging Topics

1 citations · 1 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL20251 cited

Rethinking Memory in LLM based Agents: Representations, Operations, and Emerging Topics

Yiming Du, Wenyu Huang, Danna Zheng +5

Memory is fundamental to large language model (LLM)-based agents, but existing surveys emphasize application-level use (e.g., personalized dialogue), while overlooking the atomic o…

cs.CL2025

Long-Form Information Alignment Evaluation Beyond Atomic Facts

Danna Zheng, Mirella Lapata, Jeff Z. Pan

Information alignment evaluators are vital for various NLG evaluation tasks and trustworthy LLM deployment, reducing hallucinations and enhancing user trust. Current fine-grained m…

cs.CL2024

How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency

Danna Zheng, Mirella Lapata, Jeff Z. Pan

Large Language Models (LLMs) are increasingly explored as knowledge bases (KBs), yet current evaluation methods focus too narrowly on knowledge retention, overlooking other crucial…

cs.CL2024

TrustScore: Reference-Free Evaluation of LLM Response Trustworthiness

Danna Zheng, Danyang Liu, Mirella Lapata +1

Large Language Models (LLMs) have demonstrated impressive capabilities across various domains, prompting a surge in their practical applications. However, concerns have arisen rega…

cs.CL2024

Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning

Danna Zheng, Mirella Lapata, Jeff Z. Pan

We present Archer, a challenging bilingual text-to-SQL dataset specific to complex reasoning, including arithmetic, commonsense and hypothetical reasoning. It contains 1,042 Englis…