3 papers
cs.LG2026
Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents
Bingzhen Liu, Xiaomeng Fan, Yuwei Wu +4
Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. However, these methods struggle…
cs.RO2026
A Causality-aware Infer-diagnose-refine Framework for Test-time Modality Adaptation in VLA Models
Haoyu Zhang, Yuwei Wu, Jin Chen +6
Vision-language-action (VLA) models predict sequential actions to execute tasks specified by language instructions, conditioned on visual observations and proprioceptive states. Ho…
cs.CV2026
Reliability-Prioritized Fine-Grained Generation in Multimodal Large
Xiaomeng Fan, Wei Wu, Yuwei Wu +9
Multimodal large language models (MLLMs) are increasingly expected to generate fine-grained descriptions of visual content. However, we observe and theoretically show that generati…