Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints
Songping Peng, Zhiheng Zhang, Daojian Zeng +2
Safety alignment in Large Language Models (LLMs) remains highly fragile during fine-tuning, where even benign adaptation can degrade pre-trained refusal behaviors and enable harmfu…
cs.AI2026
Agents on a Tree: Pathwise Coordination for Multi-Objective Molecular Optimization
Jia Zhang, Tengfei Ma, Tianle Li +3
Multi-objective molecular optimization requires searching vast chemical spaces under conflicting objectives, where early design decisions strongly constrain downstream outcomes. Ex…