From the 1 of 6 linked papers with an AI index.
6 papers
SAMark: A Self-Anchored Text Watermarking with Paragraph-Level Paraphrase Robustness
Jiahao Huo, Wenjie Qu, Yibo Yan +5
The paper introduces SAMark, a text watermarking method that remains detectable even after paragraph‑level paraphrasing by removing reliance on sentence order and using a hyperboli…
FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents
Haoxuan Jia, Yang Liu, Bin Chong +10
Finance LLM agents must simultaneously block prompt-induced unauthorized actions and approve legitimate multi-step business workflows. However, boundary filters often miss irrevers…
ExTax: Explainable Disinformation Detection via Persuasion, Emotion, and Narrative Role Taxonomies
Shang Luo, Yingguang Yang, Zhenchen Sun +8
The democratization of LLMs has accelerated the generation and circulation of highly fluent disinformation, making traditional syntax-semantic verification increasingly insufficien…
Why LLMs Hallucinate on Structured Knowledge: A Mechanistic Analysis of Reasoning over Linearized Representations
Shanghao Li, Jinda Han, Yibo Wang +5
In many reasoning tasks, large language models (LLMs) rely on structured external knowledge, such as graphs and tables, which is typically linearized into sequential token represen…
MultiFileTest: A Multi-File-Level LLM Unit Test Generation Benchmark and Impact of Error Fixing Mechanisms
Yibo Wang, Congying Xia, Wenting Zhao +5
Unit test generation has become a promising and important Large Language Model (LLM) use case. However, existing evaluation benchmarks for LLM unit test generation focus on functio…
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
Shanghao Li, Jinda Han, Yibo Wang +5
Graph-based Retrieval-Augmented Generation (GraphRAG) enhances Large Language Models (LLMs) by incorporating external knowledge from linearized subgraphs retrieved from knowledge g…