4 papers
MANCE: Manifold Aware Concept Erasure
Matan Avitan, Yoav Goldberg, Yanai Elazar
Concept erasure aims to remove a target concept from a representation while preserving the other information encoded in it. This is difficult because representations encode many co…
Power-Softmax: Towards Secure LLM Inference over Encrypted Data
Itamar Zimerman, Allon Adir, Ehud Aharoni +7
Modern cryptographic methods for implementing privacy-preserving LLMs such as \gls{HE} require the LLMs to have a polynomial form. Forming such a representation is challenging beca…
Efficient Decoding Methods for Language Models on Encrypted Data
Matan Avitan, Moran Baruch, Nir Drucker +2
Large language models (LLMs) power modern AI applications, but processing sensitive data on untrusted servers raises privacy concerns. Homomorphic encryption (HE) enables computati…
A Practical Method for Generating String Counterfactuals
Matan Avitan, Ryan Cotterell, Yoav Goldberg +1
Interventions targeting the representation space of language models (LMs) have emerged as an effective means to influence model behavior. Such methods are employed, for example, to…