most citedAgent AI: Surveying the Horizons of Multimodal Interaction

48 citations · 50 across the 2 of their papers we have counts for

collaborators

6 papers

cs.CV2024

OccFusion: Rendering Occluded Humans with Generative Diffusion Priors

Adam Sun, Tiange Xiang, Scott Delp +2

Most existing human rendering methods require every part of the human to be fully visible throughout the input video. However, this assumption does not hold in real-life settings w…

cs.CV2024

Few-Shot Classification of Interactive Activities of Daily Living (InteractADL)

Zane Durante, Robathan Harries, Edward Vendrow +5

Understanding Activities of Daily Living (ADLs) is a crucial step for different applications including assistive robots, smart homes, and healthcare. However, to date, few benchmar…

cs.AI2024

An Interactive Agent Foundation Model

Zane Durante, Bidipta Sarkar, Ran Gong +17

The development of artificial intelligence systems is transitioning from creating static, task-specific models to dynamic, agent-based systems capable of performing well in a wide…

cs.AI202448 cited

Agent AI: Surveying the Horizons of Multimodal Interaction

Zane Durante, Qiuyuan Huang, Naoki Wake +11

Multi-modal AI systems will likely become a ubiquitous presence in our everyday lives. A promising approach to making these systems more interactive is to embody them as agents wit…

cs.CV2024

Wild2Avatar: Rendering Humans Behind Occlusions

Tiange Xiang, Adam Sun, Scott Delp +3

Rendering the visual appearance of moving humans from occluded monocular videos is a challenging task. Most existing research renders 3D humans under ideal conditions, requiring a…

cs.AI20232 cited

MindAgent: Emergent Gaming Interaction

Ran Gong, Qiuyuan Huang, Xiaojian Ma +8

Large Language Models (LLMs) have the capacity of performing complex scheduling in a multi-agent system and can coordinate these agents into completing sophisticated tasks that req…