19 papers
AccioScene: Compositional 3D Scene Generation via Graph Diffusion and Interaction-driven Critics
Yao Wei, Matteo Toso, Pietro Morerio +3
This paper presents a framework for generating 3D indoor scenes from text prompts. Existing methods often formulate scene synthesis as an object layout prediction problem condition…
Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning
Tommaso Galliena, Stefano Rosa, Tommaso Apicella +3
Vision-Language Models (VLMs) often yield inconsistent descriptions of the same object across viewpoints, hindering the ability of embodied agents to construct consistent semantic…
Directed Semi-Simplicial Learning with Applications to Brain Activity Decoding
Manuel Lecha, Andrea Cavallo, Francesca Dominici +5
Graph Neural Networks (GNNs) excel at learning from pairwise interactions but often overlook multi-way and hierarchical relationships. Topological Deep Learning (TDL) addresses thi…
3DSGrasp: 3D Shape-Completion for Robotic Grasp
Seyed S. Mohammadi, Nuno F. Duarte, Dimitris Dimou +8
Real-world robotic grasping can be done robustly if a complete 3D Point Cloud Data (PCD) of an object is available. However, in practice, PCDs are often incomplete when objects are…
CloseUpAvatar: High-Fidelity Animatable Full-Body Avatars with Mixture of Multi-Scale Textures
David Svitov, Pietro Morerio, Lourdes Agapito +1
We present a CloseUpAvatar - a novel approach for articulated human avatar representation dealing with more general camera motions, while preserving rendering quality for close-up…
E-M3RF: An Equivariant Multimodal 3D Re-assembly Framework
Adeela Islam, Stefano Fiorini, Manuel Lecha +4
3D reassembly is a fundamental geometric problem, and in recent years it has increasingly been challenged by deep learning methods rather than classical optimization. While learnin…