1 citations · 1 across the 3 of their papers we have counts for
4 papers · 1 filter
A Single Image and Multimodality Is All You Need for Novel View Synthesis
Amirhosein Javadi, Chi-Shiang Gau, Konstantinos D. Polyzos +1
Diffusion-based approaches have recently demonstrated strong performance for single-image novel view synthesis by conditioning generative models on geometry inferred from monocular…
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
Gerasimos Chatzoudis, Konstantinos D. Polyzos, Zhuowei Li +4
Understanding the internal activations of Vision Transformers (ViTs) is critical for building interpretable and trustworthy models. While Sparse Autoencoders (SAEs) have been used…
3D Scene Rendering with Multimodal Gaussian Splatting
Chi-Shiang Gau, Konstantinos D. Polyzos, Athanasios Bacharis +2
3D scene reconstruction and rendering are core tasks in computer vision, with applications spanning industrial monitoring, robotics, and autonomous driving. Recent advances in 3D G…
ActiveInitSplat: How Active Image Selection Helps Gaussian Splatting
Konstantinos D. Polyzos, Athanasios Bacharis, Saketh Madhuvarasu +2
Gaussian splatting (GS) along with its extensions and variants provides outstanding performance in real-time scene rendering while meeting reduced storage demands and computational…