2 citations · 2 across the 3 of their papers we have counts for
4 papers · 1 filter
BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents
Jinge Wu, Hongjian Zhou, Mingde Zeng +8
Reproducing and comparing deep research agents today is hard: the same backbone evaluated on the same benchmark can report different accuracies across papers because the harness an…
A Regime Theory of Controller Class Selection for LLM Action Decisions
Zhaoyang Jiang, Zhizhong Fu, Yunsoo Kim +4
Deployed language and vision-language models must decide, on each input, whether to answer directly, retrieve evidence, defer to a stronger model, or abstain. Contrary to the commo…
Graph-based LLM over Semi-Structured Population Data for Dynamic Policy Response
Daqian Shi, Xiaolei Diao, Jinge Wu +4
Timely and accurate analysis of population-level data is crucial for effective decision-making during public health emergencies such as the COVID-19 pandemic. However, the massive…
Towards deployment-centric multimodal AI beyond vision and language
Xianyuan Liu, Jiayang Zhang, Shuo Zhou +45
Multimodal artificial intelligence (AI) integrates diverse types of data via machine learning to improve understanding, prediction, and decision-making across disciplines such as h…