2 papers
cs.CV2026
LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents
Zijian Wang, Junnan Zhu, Rongzhen Li +7
Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video tool-use agents address this ch…
cs.AI2026
Lung-R1: A Knowledge Graph-Guided LLM for Pulmonary Diagnostic Reasoning
Haoyang Zeng, Yuanxi Fu, Rongzhen Li +11
Diagnosing pulmonary diseases requires integrating heterogeneous evidence amid phenotypic variability and cross-disease overlap. Although large language models (LLMs) have shown pr…