1 paper · 1 filter
Arseniy Varlamov, Rishat Zinnatullin, Elisei Rykov +2
Tool-augmented LLMs must arbitrate between two fallible sources when a tool return conflicts with their parametric memory, yet existing evaluations measure source preference withou…