most citedGLS-CSC: A Simple but Effective Strategy to Mitigate Chinese STM Models' Over-Reliance on Superficial Clue

1 citations · 2 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CL2024

MoGU: A Framework for Enhancing Safety of Open-Sourced LLMs While Preserving Their Usability

Yanrui Du, Sendong Zhao, Danyang Zhao +6

Large Language Models (LLMs) are increasingly deployed in various applications. As their usage grows, concerns regarding their safety are rising, especially in maintaining harmless…

cs.CL2024

AS-ES Learning: Towards Efficient CoT Learning in Small Models

Nuwa Xi, Yuhan Chen, Sendong Zhao +3

Chain-of-Thought (CoT) serves as a critical emerging ability in LLMs, especially when it comes to logical reasoning. Attempts have been made to induce such ability in small models…

cs.CL2023

Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Yanrui Du, Sendong Zhao, Ming Ma +2

Extensive work has been devoted to improving the safety mechanism of Large Language Models (LLMs). However, LLMs still tend to generate harmful responses when faced with malicious…

cs.CL20231 cited

Make Your Decision Convincing! A Unified Two-Stage Framework: Self-Attribution and Decision-Making

Yanrui Du, Sendong Zhao, Haochun Wang +5

Explaining black-box model behavior with natural language has achieved impressive results in various NLP tasks. Recent research has explored the utilization of subsequences from th…

cs.CL2023

Knowledge-tuning Large Language Models with Structured Medical Knowledge Bases for Reliable Response Generation in Chinese

Haochun Wang, Sendong Zhao, Zewen Qiang +9

Large Language Models (LLMs) have demonstrated remarkable success in diverse natural language processing (NLP) tasks in general domains. However, LLMs sometimes generate responses…

cs.CL20231 cited

GLS-CSC: A Simple but Effective Strategy to Mitigate Chinese STM Models' Over-Reliance on Superficial Clue

Yanrui Du, Sendong Zhao, Yuhan Chen +5

Pre-trained models have achieved success in Chinese Short Text Matching (STM) tasks, but they often rely on superficial clues, leading to a lack of robust predictions. To address t…