Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning
Ke Xu, Yuhao Wang, Ziyang Cheng +3
Multi-hop audio-visual reasoning remains challenging for Omni-LLMs, as relevant evidence is often sparse, temporally dispersed, and distributed across both audio and visual streams…
cs.AI2026
GenTac: Generative Modeling and Forecasting of Soccer Tactics
Jiayuan Rao, Tianlin Gui, Haoning Wu +2
Modeling open-play soccer tactics is a formidable challenge due to the stochastic, multi-agent nature of the game. Existing computational approaches typically produce single, deter…