2 papers
cs.HC2026
When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks
Jieyu Zhou, Aryan Roy, Sneh Gupta +2
Existing AI agents typically execute multi-step tasks autonomously and only allow user confirmation at the end. During execution, users have little control, making the confirm-at-e…
cs.CL2026
Grounded Concreteness: Human-Like Concreteness Sensitivity in Vision-Language Models
Aryan Roy, Zekun Wang, Christopher J. MacLellan
Do vision--language models (VLMs) develop more human-like sensitivity to linguistic concreteness than text-only large language models (LLMs) when both are evaluated with text-only…