3 papers
cs.CE2026
BioXArena: Benchmarking LLM Agents on Multi-Modal Biomedical Machine Learning Tasks
Loka Li, Duzhen Zhang, Xingbo Du +11
Large language model (LLM) agents are increasingly capable of automating components of machine learning development, yet existing biomedical benchmarks mainly focus on question ans…
cs.CV2025
MindVL: Towards Efficient and Effective Training of Multimodal Large Language Models on Ascend NPUs
Feilong Chen, Yijiang Liu, Yi Huang +5
We propose MindVL, a multimodal large language model (MLLMs) trained on Ascend NPUs. The training of state-of-the-art MLLMs is often confined to a limited set of hardware platforms…
cs.CV2025
Revisiting Continual Semantic Segmentation with Pre-trained Vision Models
Duzhen Zhang, Yong Ren, Wei Cong +9
Continual Semantic Segmentation (CSS) seeks to incrementally learn to segment novel classes while preserving knowledge of previously encountered ones. Recent advancements in CSS ha…