11 papers
LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior
Qinhong Zhou, Chuang Gan, Anoop Cherian
Embodied agents operating in decentralized and partially observable environments have attracted growing attention in recent years. However, existing large language model (LLM)-base…
AssemblyBench: Physics-Aware Assembly of Complex Industrial Objects
Danrui Li, Jiahao Zhang, Bernhard Egger +4
Assembling objects from parts requires understanding multimodal instructions, linking them to 3D components, and predicting physically plausible 6-DoF motions for each assembly ste…
Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models
Furkan Mumcu, Michael J. Jones, Anoop Cherian +1
Recent video anomaly detection research has expanded rapidly with an emphasis on general models of normality intended to work across many different scenes. While this focus has led…
Leveraging Multimodal LLM Descriptions of Activity for Explainable Semi-Supervised Video Anomaly Detection
Furkan Mumcu, Michael J. Jones, Anoop Cherian +1
Existing semi-supervised video anomaly detection (VAD) methods often struggle with detecting complex anomalies involving object interactions and generally lack explainability. To o…
Agentic AI-Empowered Dynamic Survey Framework
Furkan Mumcu, Lokman Bekit, Michael J. Jones +2
Survey papers play a central role in synthesizing and organizing scientific knowledge, yet they are increasingly strained by the rapid growth of research output. As new work contin…
MMHOI: Modeling Complex 3D Multi-Human Multi-Object Interactions
Kaen Kogashi, Anoop Cherian, Meng-Yu Jennifer Kuo
Real-world scenes often feature multiple humans interacting with multiple objects in ways that are causal, goal-oriented, or cooperative. Yet existing 3D human-object interaction (…