2 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2026
MM-Snowball: Evaluating and Mitigating Hallucination Snowballing in Multimodal Multi-Turn Dialogue
Yue Jiang, Xue Jiang, Lihua Zhang +6
Multimodal large language models (MLLMs) demonstrate remarkable visual understanding, yet their reliability in interactive settings is severely undermined by hallucination snowball…
cs.CV2025★ 1 cited
Transferability of Adversarial Attacks in Video-based MLLMs: A Cross-modal Image-to-Video Approach
Linhao Huang, Xue Jiang, Zhiqiang Wang +5
Video-based multimodal large language models (V-MLLMs) have shown vulnerability to adversarial examples in video-text multimodal tasks. However, the transferability of adversarial…
cs.RO2024★ 2 cited
All Robots in One: A New Standard and Unified Dataset for Versatile, General-Purpose Embodied Agents
Zhiqiang Wang, Hao Zheng, Yunshuang Nie +11
Embodied AI is transforming how AI systems interact with the physical world, yet existing datasets are inadequate for developing versatile, general-purpose agents. These limitation…