4 papers · 1 filter
An overview of 3D Vision-Language Models
Márcus Lobo, Vitor Matias, Afonso Paiva +3
Vision-Language Models (VLMs) are reshaping computer vision by aligning visual and textual embeddings, allowing models to recognize visual concepts and reason about them using natu…
3D-MRL: Nested Multimodal 3D Representations via Matryoshka Representation Learning
Márcus Lobo, Vitor Matias, Jeová Farias +1
Vision-Language Models align point clouds with image and text embeddings, enabling zero-shot recognition, retrieval, and open-vocabulary understanding of 3D shapes. Existing multim…
From Volume Rendering to 3D Gaussian Splatting: Theory and Applications
Vitor Pereira Matias, Daniel Perazzo, Vinicius Silva +4
The problem of 3D reconstruction from posed images is undergoing a fundamental transformation, driven by continuous advances in 3D Gaussian Splatting (3DGS). By modeling scenes exp…
FLOWING: Implicit Neural Flows for Structure-Preserving Morphing
Arthur Bizzi, Matias Grynberg, Vitor Matias +7
Morphing is a long-standing problem in vision and computer graphics, requiring a time-dependent warping for feature alignment and a blending for smooth interpolation. Recently, mul…