1 citations · 1 across the 13 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
GradShield: Alignment Preserving Finetuning
Zhanhao Hu, Xiao Huang, Patrick Mendoza +4
Large Language Models (LLMs) pose a significant risk of safety misalignment after finetuning, as models can be compromised by both explicitly and implicitly harmful data. Even some…
cs.CL2025
Can LLMs Ask Good Questions?
Yueheng Zhang, Xiaoyuan Liu, Yiyou Sun +5
We evaluate questions generated by large language models (LLMs) from context, comparing them to human-authored questions across six dimensions: question type, question length, cont…