4 citations · 4 across the 4 of their papers we have counts for
6 papers
COOPER: A Unified Model for Cooperative Perception and Reasoning in Spatial Intelligence
Zefeng Zhang, Xiangzhao Hao, Hengzhu Tang +8
Visual Spatial Reasoning is crucial for enabling Multimodal Large Language Models (MLLMs) to understand object properties and spatial relationships, yet current models still strugg…
Hyperbolic-PDE GNN: Spectral Graph Neural Networks in the Perspective of A System of Hyperbolic Partial Differential Equations
Juwei Yue, Haikuo Li, Jiawei Sheng +4
Graph neural networks (GNNs) leverage message passing mechanisms to learn the topological features of graph data. Traditional GNNs learns node features in a spatial domain unrelate…
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
Taoyu Su, Jiawei Sheng, Duohe Ma +5
Multi-Modal Entity Alignment (MMEA) aims to retrieve equivalent entities from different Multi-Modal Knowledge Graphs (MMKGs), a critical information retrieval task. Existing studie…
FARM: Frequency-Aware Model for Cross-Domain Live-Streaming Recommendation
Xiaodong Li, Ruochen Yang, Shuang Wen +10
Live-streaming services have attracted widespread popularity due to their real-time interactivity and entertainment value. Users can engage with live-streaming authors by participa…
SOTOPIA-: Dynamic Strategy Injection Learning and Social Instruction Following Evaluation for Social Agents
Wenyuan Zhang, Tianyun Liu, Mengxiao Song +2
Despite the abundance of prior social strategies possessed by humans, there remains a paucity of research dedicated to their transfer and integration into social agents. Our propos…
Exploring Preference-Guided Diffusion Model for Cross-Domain Recommendation
Xiaodong Li, Hengzhu Tang, Jiawei Sheng +5
Cross-domain recommendation (CDR) has been proven as a promising way to alleviate the cold-start issue, in which the most critical problem is how to draw an informative user repres…