benchmarking 1coordinate frame alignment 1diffusion models 1inference acceleration 1large language models 1manipulation 1multi-view generalization 1parallel generation 1robot-centric perception 1vision-language-action 1
From the 2 of 61 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails
Seungbin Yang, Chaewoon Ki, Dohyun Lee +2
Recent advances in large language models have enabled web agents to autonomously execute complex tasks. In practice, users frequently provide underspecified instructions, requiring…
cs.AI2026
VisualScratchpad: Inference-time Visual Concepts Analysis in Vision Language Models
Hyesu Lim, Jinho Choi, Taekyung Kim +3
High-performing vision language models still produce incorrect answers, yet their failure modes are often difficult to explain. To make model internals more accessible and enable s…