2 papers
cs.LG2026
Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior
Christopher J. Anders, Henrique Da Silva Gameiro, Nico Daheim +1
One way to understand LLM behavior is to trace its output back to the training data. Two types of measures are commonly used for output tracing: data-similarity and data-influence.…
cs.CL2024
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
Henrique Da Silva Gameiro, Andrei Kucharavy, Ljiljana Dolamic
With the emergence of widely available powerful LLMs, disinformation generated by large Language Models (LLMs) has become a major concern. Historically, LLM detectors have been tou…