2 papers
cs.CL2026
Beyond Tokens: Concept-Level Training Objectives for LLMs
Laya Iyer, Pranav Somani, Alice Guo +2
The next-token prediction (NTP) objective has been foundational in the development of modern large language models (LLMs), driving advances in fluency and generalization. However,…
cs.CL2022
Using Natural Sentences for Understanding Biases in Language Models
Sarah Alnegheimish, Alicia Guo, Yi Sun
Evaluation of biases in language models is often limited to synthetically generated datasets. This dependence traces back to the need for a prompt-style dataset to trigger specific…