collaborators
Showing cs.CLShow all

7 papers · 1 filter

cs.CL2026

Hindsight Memory-PRM: Supervising Memory Management with Auditable Hindsight Credit

Haoxuan Jia, Yang Liu, Yingguang Yang +14

Memory operations of long-horizon LLM agents are hard to supervise: an operation's value is unobservable when it is taken. But they are special -- they leave machine-readable evide…

cs.CL2026

Characterizing Rhetorical Misalignment in Decision-Making with Language Models

Zirui Cheng, Joey Chan, Simo Du +3

Human decision-making is often shaped by a range of well-documented cognitive biases. As large language models (LLMs) become increasingly integrated into high-stakes human-AI decis…

cs.CL2026

FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

Haoxuan Jia, Yang Liu, Bin Chong +10

Finance LLM agents must simultaneously block prompt-induced unauthorized actions and approve legitimate multi-step business workflows. However, boundary filters often miss irrevers…

cs.CL2026

ExTax: Explainable Disinformation Detection via Persuasion, Emotion, and Narrative Role Taxonomies

Shang Luo, Yingguang Yang, Zhenchen Sun +8

The democratization of LLMs has accelerated the generation and circulation of highly fluent disinformation, making traditional syntax-semantic verification increasingly insufficien…

cs.CL2024

Source-Aware Training Enables Knowledge Attribution in Language Models

Muhammad Khalifa, David Wadden, Emma Strubell +4

Large language models (LLMs) learn a vast amount of knowledge during pretraining, but they are often oblivious to the source(s) of such knowledge. We investigate the problem of int…

cs.CL2023

Catwalk: A Unified Language Model Evaluation Framework for Many Datasets

Dirk Groeneveld, Anas Awadalla, Iz Beltagy +7

The success of large language models has shifted the evaluation paradigms in natural language processing (NLP). The community's interest has drifted towards comparing NLP models ac…