2 papers
cs.CV2026
Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes
Zhaoyang Wei, Bowen Jiang, Xumeng Han +6
While Multimodal Large Language Models (MLLMs) demonstrate impressive performance in benign scenarios, their cognitive reliability deteriorates significantly in complex scenes unde…
cs.IT2026
Dependency-Aware Reliability Allocation for Open-Vocabulary Scene-Graph Semantic Packets over Latency- and Energy-Constrained Wireless Visual Uplinks
Yuli Liu, Jiacheng Ruan
Wireless visual perception uplinks increasingly transmit structured semantic packets rather than raw images or opaque feature tensors. In an open-vocabulary scene graph, an indexed…