11 citations · 13 across the 4 of their papers we have counts for
4 papers
Facilitating Pornographic Text Detection for Open-Domain Dialogue Systems via Knowledge Distillation of Large Language Models
Huachuan Qiu, Shuai Zhang, Hongliang He +2
Pornographic content occurring in human-machine interaction dialogues can cause severe side effects for users in open-domain dialogue systems. However, research on detecting pornog…
Latent Jailbreak: A Benchmark for Evaluating Text Safety and Output Robustness of Large Language Models
Huachuan Qiu, Shuai Zhang, Anqi Li +2
Considerable research efforts have been devoted to ensuring that large language models (LLMs) align with human values and generate safe text. However, an excessive focus on sensiti…
A Benchmark for Understanding Dialogue Safety in Mental Health Support
Huachuan Qiu, Tong Zhao, Anqi Li +3
Dialogue safety remains a pervasive challenge in open-domain human-machine interaction. Existing approaches propose distinctive dialogue safety taxonomies and datasets for detectin…
Instance Smoothed Contrastive Learning for Unsupervised Sentence Embedding
Hongliang He, Junlei Zhang, Zhenzhong Lan +1
Contrastive learning-based methods, such as unsup-SimCSE, have achieved state-of-the-art (SOTA) performances in learning unsupervised sentence embeddings. However, in previous stud…