ai safety 1clinical decision making 1counterfactual testing 1guardrails 1hallucinated failures 1healthcare AI 1large language models 1medical agents 1self-improving agents 1self-improving systems 1
From the 2 of 21 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
ProMMSearchAgent: A Generalizable Multimodal Search Agent Trained with Process-Oriented Rewards
Wentao Yan, Shengqin Wang, Huichi Zhou +4
Training multimodal agents via reinforcement learning for knowledge-intensive visual reasoning is fundamentally hindered by the extreme sparsity of outcome-based supervision and th…
cs.CV2026
DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents
Shengqin Wang, Wentao Yan, Huichi Zhou +4
Agentic multimodal models have garnered significant attention for their ability to leverage external tools to tackle complex tasks. However, it is observed that such agents often m…