Unsupervised deep learning for semantic segmentation of multispectral LiDAR forest point clouds
arXiv:2502.06227 · doi:10.1016/j.isprsjprs.2025.07.038
Abstract
Point clouds captured with laser scanning systems from forest environments can be utilized in a wide variety of applications within forestry and plant ecology, such as the estimation of tree stem attributes, leaf angle distribution, and above-ground biomass. However, effectively utilizing the data in such tasks requires the semantic segmentation of the data into wood and foliage points, also known as leaf-wood separation. The traditional approach to leaf-wood separation has been geometry- and radiometry-based unsupervised algorithms, which tend to perform poorly on data captured with airborne laser scanning (ALS) systems, even with a high point density. While recent machine and deep learning approaches achieve great results even on sparse point clouds, they require manually labeled training data, which is often extremely laborious to produce. Multispectral (MS) information has been demonstrated to have potential for improving the accuracy of leaf-wood separation, but quantitative assessment of its effects has been lacking. This study proposes a fully unsupervised deep learning method, GrowSP-ForMS, which is specifically designed for leaf-wood separation of high-density MS ALS point clouds and based on the GrowSP architecture. GrowSP-ForMS achieved a mean accuracy of 84.3% and a mean intersection over union (mIoU) of 69.6% on our MS test set, outperforming the unsupervised reference methods by a significant margin. When compared to supervised deep learning methods, our model performed similarly to the slightly older PointNet architecture but was outclassed by more recent approaches. Finally, two ablation studies were conducted, which demonstrated that our proposed changes increased the test set mIoU of GrowSP-ForMS by 29.4 percentage points (pp) in comparison to the original GrowSP model and that utilizing MS data improved the mIoU by 5.6 pp from the monospectral case.
30 pages, 10 figures
References in corpus (11)
- Adam: A Method for Stochastic Optimization
- PCT: Point cloud transformer
- Domain-adversarial neural networks to address the appearance variability of histopathology images
- O-CNN: Octree-based Convolutional Neural Networks for 3D Shape Analysis
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention Networks
- Point Transformer V2: Grouped Vector Attention and Partition-based Pooling
- SegmentAnyTree: A sensor and platform agnostic deep learning model for tree segmentation using laser scanning data
- FOR-instance: a UAV laser scanning benchmark dataset for semantic and instance segmentation of individual trees
- Distillation with Contrast is All You Need for Self-Supervised Point Cloud Representation Learning
- Semantic segmentation of sparse irregular point clouds for leaf/wood discrimination
- Unsupervised semantic segmentation of urban high-density multispectral point clouds