activity
20242026
collaborators

11 papers

cs.LG2026

MANCE: Manifold Aware Concept Erasure

Matan Avitan, Yoav Goldberg, Yanai Elazar

Concept erasure aims to remove a target concept from a representation while preserving the other information encoded in it. This is difficult because representations encode many co…

cs.CV2025

GRADE: Quantifying Sample Diversity in Text-to-Image Models

Royi Rassin, Aviv Slobodkin, Shauli Ravfogel +2

We introduce GRADE, an automatic method for quantifying sample diversity in text-to-image models. Our method leverages the world knowledge embedded in large language models and vis…

cs.CL2025

A Practical Method for Generating String Counterfactuals

Matan Avitan, Ryan Cotterell, Yoav Goldberg +1

Interventions targeting the representation space of language models (LMs) have emerged as an effective means to influence model behavior. Such methods are employed, for example, to…

cs.LG2024

Linear Adversarial Concept Erasure

Shauli Ravfogel, Michael Twiton, Yoav Goldberg +1

Modern neural models trained on textual data rely on pre-trained representations that emerge without direct supervision. As these representations are increasingly being used in rea…

cs.CL2024

Diversity Over Quantity: A Lesson From Few Shot Relation Classification

Amir DN Cohen, Shauli Ravfogel, Shaltiel Shmidman +1

In few-shot relation classification (FSRC), models must generalize to novel relations with only a few labeled examples. While much of the recent progress in NLP has focused on scal…

cs.CL2024

Data-driven Coreference-based Ontology Building

Shir Ashury-Tahan, Amir David Nissan Cohen, Nadav Cohen +2

While coreference resolution is traditionally used as a component in individual document understanding, in this work we take a more global view and explore what can we learn about…