attribution constraints 1context residualization 1evidence recomposition 1explainable AI 1human prior alignment 1model reliability 1multimodal large language models 1subset selection attribution 1token-level explanation 1visual attribution 1
From the 2 of 10 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
BeSafe-Bench: Unveiling Behavioral Safety Risks of Situated Agents in Functional Environments
Yuxuan Li, Yi Lin, Peng Wang +2
The rapid evolution of Large Multimodal Models (LMMs) has enabled agents to perform complex digital and physical tasks, yet their deployment as autonomous decision-makers introduce…
cs.AI2025
The Safety Challenge of World Models for Embodied AI Agents: A Review
Lorenzo Baraldi, Zifan Zeng, Chongzhe Zhang +8
The rapid progress in embodied artificial intelligence has highlighted the necessity for more advanced and integrated models that can perceive, interpret, and predict environmental…
cs.AI2024
World Models: The Safety Perspective
Zifan Zeng, Chongzhe Zhang, Feng Liu +4
With the proliferation of the Large Language Model (LLM), the concept of World Models (WM) has recently attracted a great deal of attention in the AI research community, especially…