6 citations · 6 across the 15 of their papers we have counts for
6 papers · 1 filter
Continual Safety Alignment via Gradient-Based Sample Selection
Thong Bach, Dung Nguyen, Thao Minh Le +1
Large language models require continuous adaptation to new tasks while preserving safety alignment. However, fine-tuning on even benign data often compromises safety behaviors, inc…
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
Thong Bach, Thanh Nguyen-Tang, Dung Nguyen +2
Fine-tuning Large Language Models (LLMs) for downstream tasks often compromises safety alignment, even when using parameter-efficient methods like LoRA. In this work, we uncover a…
Rethinking Deep Alignment Through The Lens Of Incomplete Learning
Thong Bach, Dung Nguyen, Thao Minh Le +1
Large language models exhibit systematic vulnerabilities to adversarial attacks despite extensive safety alignment. We provide a mechanistic analysis revealing that position-depend…
Universal Multi-Domain Translation via Diffusion Routers
Duc Kieu, Kien Do, Tuan Hoang +4
Multi-domain translation (MDT) aims to learn translations between multiple domains, yet existing approaches either require fully aligned tuples or can only handle domain pairs seen…
Revisiting the Dataset Bias Problem from a Statistical Perspective
Kien Do, Dung Nguyen, Hung Le +6
In this paper, we study the "dataset bias" problem from a statistical standpoint, and identify the main cause of the problem as the strong correlation between a class attribute u a…
GEFA: Early Fusion Approach in Drug-Target Affinity Prediction
Tri Minh Nguyen, Thin Nguyen, Thao Minh Le +1
Predicting the interaction between a compound and a target is crucial for rapid drug repurposing. Deep learning has been successfully applied in drug-target affinity (DTA) problem.…