3 citations · 3 across the 1 of their papers we have counts for
2 papers
cs.CL2025
Medical Large Language Model Benchmarks Should Prioritize Construct Validity
Ahmed Alaa, Thomas Hartvigsen, Niloufar Golchini +4
Medical large language models (LLMs) research often makes bold claims, from encoding clinical knowledge to reasoning like a physician. These claims are usually backed by evaluation…
cs.AI2025★ 3 cited
BioAgents: Democratizing Bioinformatics Analysis with Multi-Agent Systems
Nikita Mehandru, Amanda K. Hall, Olesya Melnichenko +6
Creating end-to-end bioinformatics workflows requires diverse domain expertise, which poses challenges for both junior and senior researchers as it demands a deep understanding of…