autonomous driving 1depth-aware segmentation 1efficient inference 1large language models 1latent objects 1object-centric memory 1online perception 1scene discretization 1video panoptic segmentation 1video streaming 1
From the 2 of 34 linked papers with an AI index.
1 citations · 1 across the 17 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AVA: Attentive VLM Agent for Mastering StarCraft II
Weiyu Ma, Yuqian Fu, Zecheng Zhang +2
We introduce AVACraft, a multimodal StarCraft II benchmark supporting both Multi-Agent Reinforcement Learning (MARL) and Vision-Language Model (VLM) paradigms. Unlike SMAC-family e…
cs.AI2025
MLLMs are Deeply Affected by Modality Bias
Xu Zheng, Chenfei Liao, Yuqian Fu +15
Recent advances in Multimodal Large Language Models (MLLMs) have shown promising results in integrating diverse modalities such as texts and images. MLLMs are heavily influenced by…