Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Vision Transformers with Self-Distilled Registers
Yinjie Chen, Zipeng Yan, Chong Zhou +2
Vision Transformers (ViTs) have emerged as the dominant architecture for visual processing tasks, demonstrating excellent scalability with increased training data and model size. H…
cs.CV2024
PhyRecon: Physically Plausible Neural Scene Reconstruction
Junfeng Ni, Yixin Chen, Bohan Jing +7
We address the issue of physical implausibility in multi-view neural reconstruction. While implicit representations have gained popularity in multi-view 3D reconstruction, previous…