3 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Matthew W. Kenaston, Umair Ayub, Mihir Parmar +14
Despite high performance on clinical benchmarks, large language models may reach correct conclusions through faulty reasoning, a failure mode with safety implications for oncology…