12 papers
Harnessing agent memory to build lifelong AI partners for materials scientists
Siyu Liu, Bo Hu, Beilin Ye +3
Materials research advances through accumulated experience - scripts that work, protocols that are trusted, warnings attached to failed calculations or experiments, and judgement t…
RECODE: Reasoning Through Code Generation for Visual Question Answering
Junhong Shen, Mu Cai, Bo Hu +4
Multimodal Large Language Models (MLLMs) struggle with precise reasoning for structured visuals like charts and diagrams, as pixel-based perception lacks a mechanism for verificati…
A Benchmark and Knowledge-Grounded Framework for Advanced Multimodal Personalization Study
Xia Hu, Honglei Zhuang, Brian Potetz +4
The powerful reasoning of modern Vision Language Models open a new frontier for advanced personalization study. However, progress in this area is critically hampered by the lack of…
Sovereign Agents: Towards Infrastructural Sovereignty and Diffused Accountability in Decentralized AI
Botao Amber Hu, Helena Rong
AI agents deployed on decentralized infrastructures are beginning to exhibit properties that extend beyond autonomy toward what we describe as agentic sovereignty-the capacity of a…
Kami of the Commons: Towards Designing Agentic AI to Steward the Commons
Botao Amber Hu
Commons suffer from neglect, free-riding, and a persistent deficit of care. Inspired by Shinto animism -- where every forest, river, and mountain has its own \emph{kami}, a spirit…
ChartMuseum: Testing Visual Reasoning Capabilities of Large Vision-Language Models
Liyan Tang, Grace Kim, Xinyu Zhao +12
Chart understanding presents a unique challenge for large vision-language models (LVLMs), as it requires the integration of sophisticated textual and visual reasoning capabilities.…