papers

Publications (33)

cs.LG2024

Decoding-Time Language Model Alignment with Multiple Objectives

Ruizhe Shi, Yifang Chen, Yushi Hu +4

cs.CL2024

A Taxonomy of Ambiguity Types for NLP

Margaret Y. Li, Alisa Liu, Zhaofeng Wu +1

cs.CL2025

NVIDIA Nemotron 3: Efficient and Open Intelligence

NVIDIA, :, Aaron Blakeman +356

cs.CL2023

How Language Model Hallucinations Can Snowball

Muru Zhang, Ofir Press, William Merrill +2

cs.CL2026

Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

NVIDIA, :, Aaron Blakeman +571

cs.CL2026

Sampling from Your Language Model One Byte at a Time

Jonathan Hayase, Alisa Liu, Noah A. Smith +1

cs.CL2022

WANLI: Worker and AI Collaboration for Natural Language Inference Dataset Creation

Alisa Liu, Swabha Swayamdipta, Noah A. Smith +1

cs.SD2020

Bach or Mock? A Grading Function for Chorales in the Style of J.S. Bach

Alexander Fang, Alisa Liu, Prem Seetharaman +1

cs.CL2025

Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models

Hila Gonen, Terra Blevins, Alisa Liu +2

cs.CL2026

Are you going to finish that? A Practical Study of the Partial Token Problem

Hao Xu, Alisa Liu, Jonathan Hayase +2

cs.CL2025

When One LLM Drools, Multi-LLM Collaboration Rules

Shangbin Feng, Wenxuan Ding, Alisa Liu +10

cs.CL2024

Tuning Language Models by Proxy

Alisa Liu, Xiaochuang Han, Yizhong Wang +3

cs.CL2026

Are Language Models Sensitive to Morally Irrelevant Distractors?

Andrew Shaw, Christina Hahn, Catherine Rasgaitis +5

cs.CL2023

Detoxifying Text with MaRCo: Controllable Revision with Experts and Anti-Experts

Skyler Hallinan, Alisa Liu, Yejin Choi +1

cs.CL2023

Self-Instruct: Aligning Language Models with Self-Generated Instructions

Yizhong Wang, Yeganeh Kordi, Swaroop Mishra +4

cs.CL2025

Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Nathan Lambert, Jacob Morrison, Valentina Pyatkin +20

cs.LG2026

Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

NVIDIA, :, Aakshita Chandiramani +544

cs.SD2020

Incorporating Music Knowledge in Continual Dataset Augmentation for Music Generation

Alisa Liu, Alexander Fang, Gaëtan Hadjeres +2

cs.CL2024

Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data?

Jonathan Hayase, Alisa Liu, Yejin Choi +2

cs.CL2025

Nemotron 3 Nano: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

NVIDIA, :, Aaron Blakeman +311

cs.LG2025

LLAMAPIE: Proactive In-Ear Conversation Assistants

Tuochao Chen, Nicholas Batchelder, Alisa Liu +2

cs.CL2024

Inverse Scaling: When Bigger Isn't Better

Ian R. McKenzie, Alexander Lyzhov, Michael Pieler +24

cs.CL2023

That was the last straw, we need more: Are Translation Systems Sensitive to Disambiguating Context?

Jaechan Lee, Alisa Liu, Orevaoghene Ahia +2

cs.CL2025

SuperBPE: Space Travel for Language Models

Alisa Liu, Jonathan Hayase, Valentin Hofmann +3

cs.CL2026

Olmo 3

Team Olmo, :, Allyson Ettinger +66

eess.AS2020

Model selection for deep audio source separation via clustering analysis

Alisa Liu, Prem Seetharaman, Bryan Pardo

cs.CL2019

CODAH: An Adversarially Authored Question-Answer Dataset for Common Sense

Michael Chen, Mike D'Arcy, Alisa Liu +2

cs.CL2021

DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts

Alisa Liu, Maarten Sap, Ximing Lu +4

cs.CL2019

Multi-sense Definition Modeling using Word Sense Decompositions

Ruimin Zhu, Thanapon Noraset, Alisa Liu +2

cs.CL2026

Compute Optimal Tokenization

Tomasz Limisiewicz, Artidoro Pagnoni, Srini Iyer +6

cs.CL2022

Generated Knowledge Prompting for Commonsense Reasoning

Jiacheng Liu, Alisa Liu, Ximing Lu +5

cs.CL2026

Broken Tokens? Your Language Model can Secretly Handle Non-Canonical Tokenizations

Brian Siyuan Zheng, Alisa Liu, Orevaoghene Ahia +3

cs.CL2023

We're Afraid Language Models Aren't Modeling Ambiguity

Alisa Liu, Zhaofeng Wu, Julian Michael +6