Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
M3: 3D-Spatial MultiModal Memory
Xueyan Zou, Yuchen Song, Ri-Zhao Qiu +4
We present 3D Spatial MultiModal Memory (M3), a multimodal memory system designed to retain information about medium-sized static scenes through video sources for visual perception…
cs.CV2024
MVDream: Multi-view Diffusion for 3D Generation
Yichun Shi, Peng Wang, Jianglong Ye +3
We introduce MVDream, a diffusion model that is able to generate consistent multi-view images from a given text prompt. Learning from both 2D and 3D data, a multi-view diffusion mo…