Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Scaling LLaNA: Advancing NeRF-Language Understanding Through Large-Scale Training
Andrea Amaduzzi, Pierluigi Zama Ramirez, Giuseppe Lisanti +2
Recent advances in Multimodal Large Language Models (MLLMs) have shown remarkable capabilities in understanding both images and 3D data, yet these modalities face inherent limitati…
cs.CV2024
LLaNA: Large Language and NeRF Assistant
Andrea Amaduzzi, Pierluigi Zama Ramirez, Giuseppe Lisanti +2
Multimodal Large Language Models (MLLMs) have demonstrated an excellent understanding of images and 3D data. However, both modalities have shortcomings in holistically capturing th…
cs.CV2023
Looking at words and points with attention: a benchmark for text-to-shape coherence
Andrea Amaduzzi, Giuseppe Lisanti, Samuele Salti +1
While text-conditional 3D object generation and manipulation have seen rapid progress, the evaluation of coherence between generated 3D shapes and input textual descriptions lacks…