1 citations · 2 across the 6 of their papers we have counts for
6 papers
WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
Liwei Jiang, Kavel Rao, Seungju Han +8
We introduce WildTeaming, an automatic LLM safety red-teaming framework that mines in-the-wild user-chatbot interactions to discover 5.7K unique clusters of novel jailbreak tactics…
Reading Books is Great, But Not if You Are Driving! Visually Grounded Reasoning about Defeasible Commonsense Norms
Seungju Han, Junhyeok Kim, Jack Hessel +5
Commonsense norms are defeasible by context: reading books is usually great, but not when driving a car. While contexts can be explicitly described in language, in embodied scenari…
Towards Accurate Facial Landmark Detection via Cascaded Transformers
Hui Li, Zidong Guo, Seon-Min Rhee +2
Accurate facial landmarks are essential prerequisites for many tasks related to human faces. In this paper, an accurate facial landmark detector is proposed based on cascaded trans…
Pushing the Performance Limit of Scene Text Recognizer without Human Annotation
Caiyuan Zheng, Hui Li, Seon-Min Rhee +3
Scene text recognition (STR) attracts much attention over the years because of its wide application. Most methods train STR model in a fully supervised manner which requires large…
Understanding and Improving the Exemplar-based Generation for Open-domain Conversation
Seungju Han, Beomsu Kim, Seokjun Seo +2
Exemplar-based generative models for open-domain conversation produce responses based on the exemplars provided by the retriever, taking advantage of generative models and retrieva…
Self-Reorganizing and Rejuvenating CNNs for Increasing Model Capacity Utilization
Wissam J. Baddar, Seungju Han, Seonmin Rhee +1
In this paper, we propose self-reorganizing and rejuvenating convolutional neural networks; a biologically inspired method for improving the computational resource utilization of n…