16 papers
PlantRig - From Bones to Branches: Adaptation of Autoregressive Rigging Models for Plant Skeletal Reconstruction
Nathan Hu, Yang Yang, Fumio Okura
Autoregressive rigging models such as UniRig and SkinTokens perform well on articulated characters, but their ability to generalize to plant structures remains largely unexplored,…
PlantPose: Universal Plant Skeleton Estimation via Tree-constrained Graph Generation
Xinpeng Liu, Hiroaki Santo, Yosuke Toda +1
Accurate estimation of plant skeletal structures (e.g., branching structures) from images is essential for smart agriculture and plant science. Unlike human skeletons with fixed to…
Unsupervised 3D Human Pose Estimation via Conditional Multi-view Ancestral Sampling
Ryohei Goto, Takuya Fujihashi, Shunsuke Saruwatari +1
We propose a method of estimating a 3D human pose from a single view without 3D supervision. The key to our method is to leverage the 2D diffusion priors of motion diffusion models…
DP-SfM: Dual-Pixel Structure-from-Motion without Scale Ambiguity
Lilika Makabe, Kohei Ashida, Hiroaki Santo +2
Multi-view 3D reconstruction, namely, structure-from-motion followed by multi-view stereo, is a fundamental component of 3D computer vision. In general, multi-view 3D reconstructio…
NRGS: Neural Regularization for Robust 3D Semantic Gaussian Splatting
Zaiyan Yang, Xinpeng Liu, Heng Guo +3
We propose a neural regularization method that refines the noisy 3D semantic field produced by lifting multi-view inconsistent 2D features, in order to obtain an accurate and robus…
BioVITA: Biological Dataset, Model, and Benchmark for Visual-Textual-Acoustic Alignment
Risa Shinoda, Kaede Shiohara, Nakamasa Inoue +3
Understanding animal species from multimodal data poses an emerging challenge at the intersection of computer vision and ecology. While recent biological models, such as BioCLIP, h…