code generation 2performance optimization 2benchmark 1benchmarking 1boundary taxonomy 1execution feedback 1gpu programming 1isolation 1LLM agents 1motivation inference 1multimodal reasoning 1parallel computing 1
From the 4 of 17 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning
Kawai Chung, Chunkit Chan, Yauwai Yim +12
The paper introduces MultivationBench, a benchmark that tests multimodal large language models on their ability to reason about evolving human motivations across sequential visual…
cs.AI2026
Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions
Huihao Jing, Wenbin Hu, Shaojin Chen +10
The paper surveys how isolating components such as user inputs, tools, execution, inter‑agent communication, and environment can improve safety of LLM‑agent systems, presenting a b…