1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Xinmeng Hou, Ziting Chang, Zhouquan Lu +5
Large language models (LLMs) fail on over one-third of multi-hop questions with counterfactual premises and remain vulnerable to adversarial prompts that trigger biased or factuall…