activity
20212026
most citedDebate Helps Supervise Unreliable Experts

8 citations · 14 across the 12 of their papers we have counts for

collaborators
Showing cs.CLShow all

14 papers · 1 filter

cs.CL2026

No Single Best Model for Diversity: Learning a Router for Sample Diversity

Yuhan Liu, Fangyuan Xu, Vishakh Padmakumar +2

When posed with prompts that permit a large number of valid answers, comprehensively generating them is the first step towards satisfying a wide range of users. In this paper, we s…

cs.CL2025

LiteraryTaste: A Preference Dataset for Creative Writing Personalization

John Joon Young Chung, Vishakh Padmakumar, Melissa Roemmele +7

People have different creative writing preferences, and large language models (LLMs) for these tasks can benefit from adapting to each user's preferences. However, these models are…

cs.CL2025

Intent-Aware Schema Generation And Refinement For Literature Review Tables

Vishakh Padmakumar, Joseph Chee Chang, Kyle Lo +2

The increasing volume of academic literature makes it essential for researchers to organize, compare, and contrast collections of documents. Large language models (LLMs) can suppor…

cs.CL2025

Principled Content Selection to Generate Diverse and Personalized Multi-Document Summaries

Vishakh Padmakumar, Zichao Wang, David Arbour +1

While large language models (LLMs) are increasingly capable of handling longer contexts, recent work has demonstrated that they exhibit the "lost in the middle" phenomenon (Liu et…

cs.CL2025

Evaluating the Diversity and Quality of LLM Generated Content

Alexander Shypula, Shuo Li, Botong Zhang +3

Recent work suggests that preference-tuning techniques -- such as Reinforcement Learning from Human Feedback (RLHF) methods like PPO and GRPO, as well as alternatives like DPO -- r…

cs.CL2025

Measuring LLM Novelty As The Frontier Of Original And High-Quality Output

Vishakh Padmakumar, Chen Yueh-Han, Jane Pan +2

As large language models (LLMs) are increasingly used for ideation and scientific discovery, it is important to evaluate their ability to generate novel output. Prior work evaluate…