298 citations · 471 across the 27 of their papers we have counts for
12 papers · 1 filter
SocioProbe: What, When, and Where Language Models Learn about Sociodemographics
Anne Lauscher, Federico Bianchi, Samuel Bowman +1
Pre-trained language models (PLMs) have outperformed other NLP models on a wide range of tasks. Opting for a more thorough understanding of their capabilities and inner workings, r…
Easily Accessible Text-to-Image Generation Amplifies Demographic Stereotypes at Large Scale
Federico Bianchi, Pratyusha Kalluri, Esin Durmus +7
Machine learning models that convert user-written text descriptions into images are now widely available online and used by millions of users to generate millions of images a day.…
"It's Not Just Hate'': A Multi-Dimensional Perspective on Detecting Harmful Speech Online
Federico Bianchi, Stefanie Anja Hills, Patricia Rossini +3
Well-annotated data is a prerequisite for good Natural Language Processing models. Too often, though, annotation decisions are governed by optimizing time or annotator agreement. W…
ProSiT! Latent Variable Discovery with PROgressive SImilarity Thresholds
Tommaso Fornaciari, Dirk Hovy, Federico Bianchi
The most common ways to explore latent document dimensions are topic models and clustering methods. However, topic models have several drawbacks: e.g., they require us to choose th…
Data-Efficient Strategies for Expanding Hate Speech Detection into Under-Resourced Languages
Paul Röttger, Debora Nozza, Federico Bianchi +1
Hate speech is a global phenomenon, but most hate speech datasets so far focus on English-language content. This hinders the development of more effective hate speech detection mod…
Is It Worth the (Environmental) Cost? Limited Evidence for Temporal Adaptation via Continuous Training
Giuseppe Attanasio, Debora Nozza, Federico Bianchi +1
Language is constantly changing and evolving, leaving language models to become quickly outdated. Consequently, we should continuously update our models with new data to expose the…