5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2025
MTBBench: A Multimodal Sequential Clinical Decision-Making Benchmark in Oncology
Kiril Vasilev, Alexandre Misrahi, Eeshaan Jain +5
Multimodal Large Language Models (LLMs) hold promise for biomedical reasoning, but current benchmarks fail to capture the complexity of real-world clinical workflows. Existing eval…
cs.AI2025★ 5 cited
AgentRxiv: Towards Collaborative Autonomous Research
Samuel Schmidgall, Michael Moor
Progress in scientific discovery is rarely the result of a single "Eureka" moment, but is rather the product of hundreds of scientists incrementally working together toward a commo…