6 citations · 9 across the 3 of their papers we have counts for
3 papers
cs.AI2024
GrounDial: Human-norm Grounded Safe Dialog Response Generation
Siwon Kim, Shuyang Dai, Mohammad Kachuee +3
Current conversational AI systems based on large language models (LLMs) are known to generate unsafe responses, agreeing to offensive user input or including toxic content. Previou…
cs.LG2023★ 6 cited
Probabilistic Concept Bottleneck Models
Eunji Kim, Dahuin Jung, Sangha Park +2
Interpretable models are designed to make decisions in a human-interpretable manner. Representatively, Concept Bottleneck Models (CBM) follow a two-step process of concept predicti…
cs.LG2023★ 3 cited
On the Impact of Knowledge Distillation for Model Interpretability
Hyeongrok Han, Siwon Kim, Hyun-Soo Choi +1
Several recent studies have elucidated why knowledge distillation (KD) improves model performance. However, few have researched the other advantages of KD in addition to its improv…