2 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.AI2023★ 2 cited
Enhancing the Spatial Awareness Capability of Multi-Modal Large Language Model
Yongqiang Zhao, Zhenyu Li, Zhi Jin +6
The Multi-Modal Large Language Model (MLLM) refers to an extension of the Large Language Model (LLM) equipped with the capability to receive and infer multi-modal data. Spatial awa…
cs.AR2023★ 2 cited
SpOctA: A 3D Sparse Convolution Accelerator with Octree-Encoding-Based Map Search and Inherent Sparsity-Aware Processing
Dongxu Lyu, Zhenyu Li, Yuzhou Chen +3
Point-cloud-based 3D perception has attracted great attention in various applications including robotics, autonomous driving and AR/VR. In particular, the 3D sparse convolution (Sp…
cs.CV2021★ 2 cited
ForgeryNet -- Face Forgery Analysis Challenge 2021: Methods and Results
Yinan He, Lu Sheng, Jing Shao +19
The rapid progress of photorealistic synthesis techniques has reached a critical point where the boundary between real and manipulated images starts to blur. Recently, a mega-scale…