3 papers
cs.CV2025
Selftok: Discrete Visual Tokens of Autoregression, by Diffusion, and for Reasoning
Bohan Wang, Zhongqi Yue, Fengda Zhang +15
We completely discard the conventional spatial prior in image representation and introduce a novel discrete visual tokenizer: Self-consistency Tokenizer (Selftok). At its design co…
cs.CV2024
Fine Structure-Aware Sampling: A New Sampling Training Scheme for Pixel-Aligned Implicit Models in Single-View Human Reconstruction
Kennard Yanting Chan, Fayao Liu, Guosheng Lin +2
Pixel-aligned implicit models, such as PIFu, PIFuHD, and ICON, are used for single-view clothed human reconstruction. These models need to be trained using a sampling training sche…
cs.CV2024
IntegratedPIFu: Integrated Pixel Aligned Implicit Function for Single-view Human Reconstruction
Kennard Yanting Chan, Guosheng Lin, Haiyu Zhao +1
We propose IntegratedPIFu, a new pixel aligned implicit model that builds on the foundation set by PIFuHD. IntegratedPIFu shows how depth and human parsing information can be predi…