1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2026
EVA01: Unified Native 3D Understanding and Generation via Mixture-of-Transformers
Zongyuan Yang, Mingjing Yi, Wanli Ma +8
This paper addresses the challenge of integrating 3D meshes as a native modality within Multimodal Large Language Models (MLLMs). Diffusion-based large reconstruction models decoup…
q-bio.NC2024★ 1 cited
Understanding Auditory Evoked Brain Signal via Physics-informed Embedding Network with Multi-Task Transformer
Wanli Ma, Xuegang Tang, Jin Gu +2
In the fields of brain-computer interaction and cognitive neuroscience, effective decoding of auditory signals from task-based functional magnetic resonance imaging (fMRI) is key t…