most citedMedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine

2 citations · 5 across the 3 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter