activity
20242026
collaborators
Showing cs.CLShow all

6 papers · 1 filter

cs.CL2026

CreativityPrism: A Cross-Domain Evaluation Framework for Large Language Model Creativity

Zhaoyi Joey Hou, Bowei Alvin Zhang, Yining Lu +9

Creativity is often seen as a hallmark of human intelligence. While large language models(LLMs) are increasingly perceived as generating creative text, there is still no cross-doma…

cs.CL2026

Uncovering Cross-Objective Interference in Multi-Objective Alignment

Yining Lu, Meng Jiang

We study a persistent failure mode in multi-objective alignment for large language models (LLMs): training improves performance on only a subset of objectives while causing others…

cs.CL2025

Optimizing Decomposition for Optimal Claim Verification

Yining Lu, Noah Ziems, Hy Dang +1

Current research on the \textit{Decompose-Then-Verify} paradigm for evaluating the factuality of long-form text typically treats decomposition and verification in isolation, overlo…

cs.CL2025

Benchmarking Language Model Creativity: A Case Study on Code Generation

Yining Lu, Dixuan Wang, Tianjian Li +4

As LLMs become increasingly prevalent, it is interesting to consider how ``creative'' these models can be. From cognitive science, creativity consists of at least two key character…

cs.CL2024

AnaloBench: Benchmarking the Identification of Abstract and Long-context Analogies

Xiao Ye, Andrew Wang, Jacob Choi +6

Humans regularly engage in analogical thinking, relating personal experiences to current situations (X is analogous to Y because of Z). Analogical thinking allows humans to solve p…

cs.CL2024

RORA: Robust Free-Text Rationale Evaluation

Zhengping Jiang, Yining Lu, Hanjie Chen +3

Free-text rationales play a pivotal role in explainable NLP, bridging the knowledge and reasoning gaps behind a model's decision-making. However, due to the diversity of potential…