◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tomer Ashuach

4 papers hereh-index 223 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL4

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.CL2026

Reasoning Models Know What's Important, and Encode It in Their Activations

Yaniv Nikankin, Martin Tutek, Tomer Ashuach +2

Language models often solve complex tasks by generating long reasoning chains, consisting of many steps with varying importance. While some steps are crucial for generating the fin…

cs.CL2026

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

Tomer Ashuach, Shai Gretz, Yoav Katz +2

Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether large language models possess si…

cs.CL2026

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

Tomer Ashuach, Dana Arad, Aaron Mueller +2

As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preserving model utility has become pa…

cs.CL2025

REVS: Unlearning Sensitive Information in Language Models via Rank Editing in the Vocabulary Space

Tomer Ashuach, Martin Tutek, Yonatan Belinkov

Language models (LMs) risk inadvertently memorizing and divulging sensitive or personally identifiable information (PII) seen in training data, causing privacy concerns. Current ap…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.