5 papers
Measuring Embedding Sensitivity to Authorial Style in French: Comparing Literary Texts with Language Model Rewritings
Benjamin Icard, Lila Sainero, Alice Breton +2
Large language models (LLMs) can convincingly imitate human writing styles, yet it remains unclear how much stylistic information is encoded in embeddings from any language model a…
From Noise to Signal: When Outliers Seed New Topics
Evangelia Zve, Gauvain Bourgne, Benjamin Icard +1
Outliers in dynamic topic modeling are typically treated as noise, yet we show that some can serve as early signals of emerging topics. We introduce a temporal taxonomy of news-doc…
From Outliers to Topics in Language Models: Anticipating Trends in News Corpora
Evangelia Zve, Benjamin Icard, Alice Breton +3
This paper examines how outliers, often dismissed as noise in topic modeling, can act as weak signals of emerging topics in dynamic news corpora. Using vector embeddings from state…
Beyond the Spell: A Dynamic Logic Analysis of Misdirection
Benjamin Icard, Raul Fervari
Misdirection can be defined as the intentional action of causing some misrepresentation in an agent, or in a group of agents. Such misrepresentations may result from verbal actions…
Embedding Style Beyond Topics: Analyzing Dispersion Effects Across Different Language Models
Benjamin Icard, Evangelia Zve, Lila Sainero +2
This paper analyzes how writing style affects the dispersion of embedding vectors across multiple, state-of-the-art language models. While early transformer models primarily aligne…