activity
20242026
collaborators

19 papers

cs.LG2026

AccioScene: Compositional 3D Scene Generation via Graph Diffusion and Interaction-driven Critics

Yao Wei, Matteo Toso, Pietro Morerio +3

This paper presents a framework for generating 3D indoor scenes from text prompts. Existing methods often formulate scene synthesis as an object layout prediction problem condition…

cs.CV2026

Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning

Tommaso Galliena, Stefano Rosa, Tommaso Apicella +3

Vision-Language Models (VLMs) often yield inconsistent descriptions of the same object across viewpoints, hindering the ability of embodied agents to construct consistent semantic…

cs.LG2026

Directed Semi-Simplicial Learning with Applications to Brain Activity Decoding

Manuel Lecha, Andrea Cavallo, Francesca Dominici +5

Graph Neural Networks (GNNs) excel at learning from pairwise interactions but often overlook multi-way and hierarchical relationships. Topological Deep Learning (TDL) addresses thi…

cs.RO2025

3DSGrasp: 3D Shape-Completion for Robotic Grasp

Seyed S. Mohammadi, Nuno F. Duarte, Dimitris Dimou +8

Real-world robotic grasping can be done robustly if a complete 3D Point Cloud Data (PCD) of an object is available. However, in practice, PCDs are often incomplete when objects are…

cs.CV2025

CloseUpAvatar: High-Fidelity Animatable Full-Body Avatars with Mixture of Multi-Scale Textures

David Svitov, Pietro Morerio, Lourdes Agapito +1

We present a CloseUpAvatar - a novel approach for articulated human avatar representation dealing with more general camera motions, while preserving rendering quality for close-up…

cs.CV2025

E-M3RF: An Equivariant Multimodal 3D Re-assembly Framework

Adeela Islam, Stefano Fiorini, Manuel Lecha +4

3D reassembly is a fundamental geometric problem, and in recent years it has increasingly been challenged by deep learning methods rather than classical optimization. While learnin…