activity
20162024
most citedSUMBT: Slot-Utterance Matching for Universal and Scalable Belief Tracking

24 citations · 64 across the 14 of their papers we have counts for

collaborators
Showing 2023Show all

7 papers · 1 filter

cs.CL2023

LifeTox: Unveiling Implicit Toxicity in Life Advice

Minbeom Kim, Jahyun Koo, Hwanhee Lee +3

As large language models become increasingly integrated into daily life, detecting implicit toxicity across diverse contexts is crucial. To this end, we introduce LifeTox, a datase…

cs.CR2023★ 21 cited

ProPILE: Probing Privacy Leakage in Large Language Models

Siwon Kim, Sangdoo Yun, Hwaran Lee +3

The rapid advancement and widespread use of large language models (LLMs) have raised significant concerns regarding the potential leakage of personally identifiable information (PI…

cs.CL2023★ 2 cited

KoBBQ: Korean Bias Benchmark for Question Answering

Jiho Jin, Jiseon Kim, Nayeon Lee +3

The Bias Benchmark for Question Answering (BBQ) is designed to evaluate social biases of language models (LMs), but it is not simple to adapt this benchmark to cultural contexts ot…

cs.CL2023★ 5 cited

KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application

Hwaran Lee, Seokhee Hong, Joonsuk Park +3

Large language models (LLMs) learn not only natural text generation abilities but also social biases against different demographic groups from real-world data. This poses a critica…

cs.CL2023

SQuARe: A Large-Scale Dataset of Sensitive Questions and Acceptable Responses Created Through Human-Machine Collaboration

Hwaran Lee, Seokhee Hong, Joonsuk Park +10

The potential social harms that large language models pose, such as generating offensive content and reinforcing biases, are steeply rising. Existing works focus on coping with thi…

cs.AI2023

Query-Efficient Black-Box Red Teaming via Bayesian Optimization

Deokjae Lee, JunYeong Lee, Jung-Woo Ha +4

The deployment of large-scale generative models is often restricted by their potential risk of causing harm to users in unpredictable ways. We focus on the problem of black-box red…