◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Roy Schwartz

5 papers hereh-index 243 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author4

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CL5
same name
  • Roy Schwartz — 5 papers, h 4
  • Roy Schwartz — 3 papers, h 2
  • Roy Schwartz — 3 papers, h 32
  • Roy Schwartz — 1 paper, h 12
  • Roy Schwartz — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.CL2026

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

Federico Marcuzzi, Xuefei Ning, Roy Schwartz +1

As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current literature suffers from wide…

cs.CL2026

Post-training is (Massive) Supervised Learning

Michael Hassid, Yossi Adi, Roy Schwartz

The prevailing paradigm for training LLMs has evolved to rely on a massive post-training phase consisting of SFT and RL. In this position paper, we argue that this methodology effe…

cs.CL2026

Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic

Yuval Reif, Guy Kaplan, Roy Schwartz

Large language models (LLMs) often encode word-form variation (e.g., walk vs. walked) as linear directions in the embedding space. However, standard tokenization algorithms treat s…

cs.CL2026

Why Fine-Tuning Encourages Hallucinations and How to Fix It

Guy Kaplan, Zorik Gekhman, Zhen Zhu +5

Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information through supervised fine-tuning…

cs.CL2025

SpeLLM: Character-Level Multi-Head Decoding

Amit Ben-Artzy, Roy Schwartz

Scaling LLM vocabulary is often used to reduce input sequence length and alleviate attention's quadratic cost. Yet, current LLM architectures impose a critical bottleneck to this p…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.