1 paper
Ernest Lim, Yajie Vera He, Jared Joselowitz +9
Despite the growing use of large language models (LLMs) in clinical dialogue systems, existing evaluations focus on task completion or fluency, offering little insight into the beh…